GPT-6.1 Sol vs Opus 5.5
Two Go interpreters pass the tests. One runs faster. One knows when to stop.

Based on a third-party coding test. Scenes and dialogue are dramatized.
MODEL COMPARISONS
Specific tasks, reported results, and a little comic relief.
Two Go interpreters pass the tests. One runs faster. One knows when to stop.

Based on a third-party coding test. Scenes and dialogue are dramatized.
Two tutors, similar teaching scores, and one awkward bill.

Based on Braintrust’s reported teaching evaluation. Costs are specific to that workload.