We have just launched our enhanced evidence-based medical model, Baichuan-M2 Plus, while simultaneously upgrading its supporting application Baixiaoying and opening the API.
LLMS
-
Testing GPT-4.1 for writing tasks
By
–
Hmm, I am currently sticking with GPT 4.1 for writing, probably I need to test it a bit more as well
-
Jensen Huang: AI Will Let Us Converse With Living Cells
By
–
« Tu pourrais parler à une cellule comme tu parles à un chatbot. »
— VISION IA (@vision_ia) 24 octobre 2025
Jensen Huang affirme que la même IA qui comprend le langage comprendra bientôt la vie elle-même transformant la biologie en quelque chose avec laquelle nous pourrons converser.
« Quelles sont tes propriétés ? À… pic.twitter.com/DGXxlT6PEZ« You could talk to a cell the way you talk to a chatbot. » Jensen Huang asserts that the same AI that understands language will soon understand life itself, transforming biology into something we can converse with. « What are your properties? What can you bind to? »
-
GLM 4.6 Remains Viable for Coding Agents Long-Term
By
–
GLM 4.6 will continue to exist and be useful for powering coding agents even if every single AI lab goes bust and vanishes
-

SnorkelSpatial Benchmark Tests LLM Spatial Reasoning Abilities
By
–
New benchmark drop SnorkelSpatial tests how well LLMs can think in space, following text-based moves and rotations in a 2D world.
-
New Annotations Feature Released for Google AI Studio
By
–
The annotations feature for @GoogleAIStudio Built was also released! For example, you can now precisely instruct Gemini on which animations to add to your app. https://t.co/AHbrkPuBnQ pic.twitter.com/LDxjYsQsBk
— 🚨 AI News | TestingCatalog (@testingcatalog) 23 octobre 2025The annotations feature for @GoogleAIStudio Built was also released! For example, you can now precisely instruct Gemini on which animations to add to your app.
-
Stitch Application Upgraded with Gemini 2.5 Pro
By
–
Stitch is now powered by max-tuned Gemini 2.5 Pro, and it is killing it. There is a visible difference from what Stitch was capable of before. https://t.co/CNSuLzOAdX pic.twitter.com/R1awz7X210
— 🚨 AI News | TestingCatalog (@testingcatalog) 23 octobre 2025Stitch is now powered by max-tuned Gemini 2.5 Pro, and it is killing it. There is a visible difference from what Stitch was capable of before.
-
Claude Opus rate limits on $20 subscription plan
By
–
I can send about 15 messages to Opus on the $20 plan before I get locked out for 4 hours.
-

Meta’s Soft Tokens Enable LLMs to Invent Recursive Representations
By
–
This paper from Meta about "Soft Tokens" in RL is interesting; it allows LLMs to invent their own non-discrete (recursive) representations in order to solve problems better… Results are mixed though: it's only a few percent better on GSM8k from pass@4 onwards, and pass@32 just
