guessed . Can you beat them?
How long will take to ? A person takes about .
Racing Opus 4.8 (Anthropic) against a typical person on write screenplay dialogue: 40 min by hand versus 18 s for the model, 134× faster, 93% accurate on this task's rubric. ✓ Good fit. Opus 4.8 handles this comfortably. It used 1.0k tokens for a cost of $0.07: see how that stacks up against every other model →
Choose a task and model. Compare its recorded or estimated speed, accuracy, and cost with a typical human time.
guessed . Can you beat them?
How long will take to ? A person takes about .
Run the same task on Claude and GPT side by side.
"You" is a typical time for a person doing the task by hand, unless you've taken the 15-second typing test above, in which case it's your own. "AI" is a recorded run for the selected model, shown to scale. Stage times, accuracy, and token cost are placeholder profiles until runs are recorded and scored. How we measure → · FAQ → · Model comparison → · Open dataset →