Racing Sonnet 5 (Anthropic) against a typical person on draft a phd research proposal: 8 hr by hand versus 35 s for the model, 823× faster, 81% accurate on this task's rubric. △ Sonnet 5 can do it, but Opus 4.8 would do it better. It used 951 tokens for a cost of $0.01 — see how that stacks up against every other model →
Choose a task you do every week, then pick an AI model. Watch how fast it finishes, and see what happened in those seconds: what it read, understood, thought through, wrote — and how accurate it was.
Run the same task on Claude and GPT side by side — with effort levels and Extended Thinking.
"You" is a typical time for a person doing the task by hand. "AI" is a recorded run for the selected model, shown to scale. Stage times, accuracy, and token cost are placeholder profiles until runs are recorded and scored. How we measure → · FAQ → · Model comparison →