Skip to benchmark contentMotionBenchHosted by Baz Studio Human assessment
July diagnostic pilot33 routes · diagnostic datasetJuly diagnostic pilot
openai · O+C
GPT-5.5 Low
gpt-5.5-lowDiagnostic tierA
Distinct stretched-form transition and clean typographic landing.
Historical diagnostic evidence; not a durable capability claim.