guessed . Can you beat them?
How long will take to ? A person takes about .
Racing Sonnet 5 (Anthropic) against a typical person on adapt a novel chapter to screenplay: 3 hr by hand versus 18 s for the model, 600× faster, 82% accurate on this task's rubric. △ Sonnet 5 can do it, but Opus 4.8 would do it better. It used 506 tokens for a cost of $0.0048: see how that stacks up against every other model →
Choose a task and model. Compare its recorded or estimated speed, accuracy, and cost with a typical human time.
guessed . Can you beat them?
How long will take to ? A person takes about .
Run the same task on Claude and GPT side by side.
"You" is a typical time for a person doing the task by hand, unless you've taken the 15-second typing test above, in which case it's your own. "AI" is a recorded run for the selected model, shown to scale. Stage times, accuracy, and token cost are placeholder profiles until runs are recorded and scored. How we measure → · FAQ → · Model comparison → · Open dataset →