01
Harness
MODEL LOCK
openai/gpt-5.6-solLane ready
Run this harness or launch the full field.LATENCY—
EVENTS—
TOKENS—
HARNESS/GYM
agent runtime comparison bench
Launch one controlled turn across six runtimes. Compare the answer, event trail, latency, usage, and failure shape without changing the task.
LIVE COMPARISON
openai/gpt-5.6-solLane ready
Run this harness or launch the full field.openai/gpt-5.4-miniLane ready
Run this harness or launch the full field.openai/gpt-5.4-miniLane ready
Run this harness or launch the full field.