t08: Write the tests -- GPT-6.1 Sol vs Claude Opus 5.5


Write a full test suite for a Python pricing engine that has none.
Opus: 169.2 s. Sol: 1023.0 s. Sol took 6.0 times as long.
Opus (s)Sol (s)
Hidden tests: Opus 9 of 9, Sol 9 of 9.
Blind judge pick: Tie.
"Both suites are exceptionally thorough, well-structured, and achieve perfect mutation scores. They both use custom assertions, parameterize effectively with subtests, and cover deep edge cases like fractional cent rounding and proportional discount allocation."
"An exceptionally thorough and well-structured test suite that covers all edge cases and successfully kills all mutants."
"An exceptionally thorough and well-structured test suite that successfully covers all edge cases and catches all hidden bugs."