Can we settle on "year of vision LMs for agents"?
LLMS
-

Gemini 2.5 Pro achieves significant LiveBench performance improvements
By
–
A mere ~16 point jump on http://
livebench.ai. Such a good model! Gemini 2.5 Pro -

ChatGPT Image Generation Upgrade and Prompt Example
By
–
BREAKING: Image generation in ChatGPT just got a HUGE upgrade! Prompt I used:
"Create image of phone photo of friends chilling at a party, taken with a film camera" -
AI Model Prompt Adherence with Large Scale Objects and Text
By
–
The prompt adherence is carzy, you can just throw an insane amount objects, and still get it right, plus handling text.
-

Meeting Satya Nadella to discuss SLLM in Korean
By
–
Super exciting to meet @satyanadella in Korean and discuss SLLM.
-
Google Releases Gemini 2.5 Pro Experimental with 1M Token Context
By
–
Google released Gemini 2.5 Pro Experimental, the first model in its Gemini 2.5 family
— Rowan Cheung (@rowancheung) 26 mars 2025
—#1 on the LMArena
—SOTA capabilities across benchmarks for coding, math, science, and more
—Visual reasoning
—1M token context window (2M coming soon!)pic.twitter.com/3G7MfXsfxkGoogle released Gemini 2.5 Pro Experimental, the first model in its Gemini 2.5 family —#1 on the LMArena
—SOTA capabilities across benchmarks for coding, math, science, and more
—Visual reasoning
—1M token context window (2M coming soon!)




