Google to Launch Gemini 3.2 Flash at I/O on May 20, Matching GPT-5.5 Performance at 1/15 Cost

AT19.17%
ON0.27%
MAY-1.21%
According to Abacus.AI CEO Bindu Reddy, Google plans to unveil Gemini 3.2 Flash during its I/O conference on May 20, with performance reaching 92% of GPT-5.5 on coding and reasoning tasks while cutting inference costs to just one-fifteenth to one-twentieth of the latter. Most queries will have latency below 200 milliseconds. Reddy attributed the breakthrough to Google's distillation and sparsity techniques, which compress a frontier model into the Flash tier without the typical performance cliff typically seen in model optimization.
Disclaimer: The information on this page may come from third-party sources and is for reference only. It does not represent the views or opinions of Gate and does not constitute any financial, investment, or legal advice. Virtual asset trading involves high risk. Please do not rely solely on the information on this page when making decisions. For details, see the Disclaimer.
Comment
0/400
No comments