
OpenAI's Efficiency Ledger: Serving Costs Down 20%, ARC-AGI-3 Up 3x With No Model Change
The "Building abundant intelligence" essay carries real engineering numbers: GPT-5.6 Sol cut serving costs 20%, speculative decoding gained 15%, and two settings moved ARC-AGI-3 from 13.3% to 38.3% with six times fewer tokens.









