Racing Sonnet 5 (Anthropic) against a typical person on design a meta-analysis plan: 4 hr by hand versus 22 s for the model, 655× faster, 79% accurate on this task's rubric. △ Sonnet 5 can do it, but Opus 4.8 would do it better. It used 610 tokens for a cost of $0.0087 — see how that stacks up against every other model →
Choose a task you do every week, then pick an AI model. Watch how fast it finishes, and see what happened in those seconds: what it read, understood, thought through, wrote — and how accurate it was.
Run the same task on Claude and GPT side by side — with effort levels and Extended Thinking.
"You" is a typical time for a person doing the task by hand. "AI" is a recorded run for the selected model, shown to scale. Stage times, accuracy, and token cost are placeholder profiles until runs are recorded and scored. How we measure → · FAQ → · Model comparison →