How I AI · Tuesday, September 22, 2026
Beyond cost cuts, the new Opus 55, GPT-6 Sole, and GPT-6 Luna models are emphasizing speed and token efficiency. This focus aims to reduce output, token usage, and ultimately, the cost for users.
“Okay so the real thing that they're focusing on not just cutting the cost but also cutting on cached inputs and just speed and token efficiency”
“and so you're gonna see both sort of like output drop token use drop and cost drop it's really nice”