The Call Center Doctors: 1-877-766-3765

t08: Write the tests -- GPT-6.1 Sol vs Claude Opus 5.5

Cartoon: the golden octopus mascotCartoon: the orange sun mascot

Write a full test suite for a Python pricing engine that has none.

Opus: 169.2 s. Sol: 1023.0 s. Sol took 6.0 times as long.

Opus (s)Sol (s)

  1. 169.21023.0

Hidden tests: Opus 9 of 9, Sol 9 of 9.

Blind judge pick: Tie.

Why the judge picked

"Both suites are exceptionally thorough, well-structured, and achieve perfect mutation scores. They both use custom assertions, parameterize effectively with subtests, and cover deep edge cases like fractional cent rounding and proportional discount allocation."

The judge on Opus

"An exceptionally thorough and well-structured test suite that covers all edge cases and successfully kills all mutants."

The judge on Sol

"An exceptionally thorough and well-structured test suite that successfully covers all edge cases and catches all hidden bugs."