AI Dynamics

Global AI News Aggregator

About

Rumored Gemini Flash achieves 92% GPT-5.5 performance at 15-20x lower inference cost

Rumors about the new Gemini Flash coming in. And holy, if true then big: 92% of GPT-5.5’s coding and reasoning performance, reportedly at 15–20x lower inference cost. And the latency? Sub-200ms for most queries. That would be nuts. no joke.

→ View original post on X — @kimmonismus