Racing GPT-5.6 Luna (OpenAI) against a typical person on write screenplay dialogue: 40 min by hand versus 2.3 s for the model, 1,067× faster, 62% accurate on this task's rubric. ✗ Too small for this one. Try GPT-5.6 Terra. It used 152 tokens for a cost of $0.0004 — see how that stacks up against every other model →
Choose a task you do every week, then pick an AI model. Watch how fast it finishes, and see what happened in those seconds: what it read, understood, thought through, wrote — and how accurate it was.
Run the same task on Claude and GPT side by side — with effort levels and Extended Thinking.
"You" is a typical time for a person doing the task by hand. "AI" is a recorded run for the selected model, shown to scale. Stage times, accuracy, and token cost are placeholder profiles until runs are recorded and scored. How we measure → · FAQ → · Model comparison →