hmm i liked the 32B variety that was released earlier need to test myself and see if its passes my vibe checks will report back on it later tonight
GENERATIVE AI
-

KAT-Dev-72B-Exp: Agentic Coding Model Ranks #2 SWE-Bench
By
–
HUGE a new Agentic coding model, fits on 4x RTX 3090s @ 4-bit, fully local KAT-Dev-72B-Exp by Kwaipilot – Claude Code setup guide included – ranks #2 on SWE-Bench Verified – excels at long-horizon coding + tool-use – multi-stage tuned: Mid-Training, SFT + RFT, Agentic RL
-

BAIR Researchers Win Outstanding Paper Award at COLM2025
By
–
Congratulations to BAIR Researchers from @trevordarrell lab whose paper "Hidden in plain sight: VLMs overlook their visual representations," was awarded an Outstanding Paper Award at #COLM2025 this week in Montreal, Canada. @xkungfu @tylerraye @databoydg
-
Unrealistic AI Hype and Follower Disappointment Cycles
By
–
Se sobre-emocionan ante la expectativa de algo que no tiene sentido, y luego se sobre-decepcionan cuando esas expectativas falsas no se cumplen. Y con ellos sus seguidores.
-

AI Influencers Misunderstand METR Benchmark Results for Sonnet
By
–
The number of AI influencers who are surprised that Sonnet 4.5 didn't achieve a better position on the METR benchmark, when they were saying it "could work autonomously for 30 hours," worries me. They're not understanding anything about what these benchmarks measure. They're
-
DeepThink IMO Lite Now Available Public Gemini Ultra
By
–
Correction: it's a public model access through Gemini Ultra subscription (not API). I guess the main point is DeepThink IMO lite is available for the public to assess 🙂
-
Gemini Ultra Subscription Now Offers Public Model Access
By
–
Yea, sorry publicly available model access through Gemini Ultra subscription
-
DeepAgent desktop feature coming to ChatLLM
By
–
It’s there in DeepAgent desktop – will be coming to ChatLLM as well
-
Future where AI videos are presumed authentic
By
–
Or, the opposite might happen. All compelling video footage (defamatory or otherwise) will be assumed to be AI by default.
-

Gemini 2.5 Deep Think Achieves State-of-the-Art on FrontierMath
By
–
Gemini 2.5 Deep Think is SoTA on FrontierMath!