Congrats to the @MiniMax_AI team on the release of MiniMax M3, a long-context multimodal model for text, image, and video reasoning. Try it today with our free GPU-accelerated endpoint on http://
build.nvidia.com. Details: https://
nvda.ws/4v4BWhD
RESEARCH
-

Nvidia AI congratulates MiniMax on M3 multimodal model release
By
–
-

Frontier LLMs outperform clinical AI tools in medical evaluations
By
–
There has been a push to use OpenEvidence AI for doctors. But this paper suggests general models are much better: “Frontier LLMs outperformed clinical AI tools in all three evaluations. Clinical AI tools performed comparably to auto-enabled Google Search AI Overview on the RCQ.”
-
Proposal to evaluate Omni and build a routing system
By
–
Yes maybe we should evaluate Omni @victormustar or https://openrouter.ai/docs/guides/routing/provider-selection … @alexatallah? Or does someone want to build a routing system and evaluate on AA?
-
Machine Learning with Amazon SageMaker for Big Data and AI
By
–
Machine Learning with Amazon SageMaker! #BigData #Analytics #DataScience #AI #MachineLearning #IoT #IIoT #PyTorch #Python #RStats #TensorFlow #Java #ReactJS #GoLang #CloudComputing #Serverless #DataScientist #Linux #Programming #Coding #100DaysofCode https://
geni.us/SageMaker-AWS -

General AI models outperform specialized medical sources in study
By
–
For medical information, general AI frontier models (Google, OpenAI, Anthropic) outperformed specialized @EvidenceOpen and @UpToDate as assessed by 12 US clinicians, randomized and blinded to which model and extensive testing/benchmarks. This was not anticipated.
-

AI is reshaping discovery in mathematics and physics
By
–
How #AI is reshaping discovery in maths and physics
by Mikhail Burtsev Yang-Hui He @Nature Learn more: https://
bit.ly/4vCrk9O #ArtificialIntelligence #MachineLearning #ML -

AI Agents Learn from Failures and Evolve: New Survey
By
–
What if AI agents could not only collaborate but also learn from their own failures and evolve? Researchers from Xi'an Jiaotong University, Lenovo, and the University of Sydney present a new survey. They introduce the LIFE progression: build agent capabilities, integrate them
-
Techniques for randomness and diversity in language model outputs
By
–
Summary of things: – turn up randomness and ban the most likely words (temperature + min-p + XTC sampling)
– ask for several different options at once, seed each with random constraints (verbalized sampling + entropy injection)
– give it a memory of what it's said and pick the -

Carnegie Mellon benchmark exposes safety risks of AI coding agents
By
–
A new benchmark just exposed the dirty secret behind every coding agent. Millions of developers now let AI agents write entire features unsupervised. A Carnegie Mellon paper tested whether that code is safe to ship. The team built SusVibes, a benchmark of 200 real coding
-
AI context awareness merges multiple sources into one query
By
–
The bigger unlock is context awareness. You can bring:
→ Multiple tabs
→ Documents
→ Images Into one query. That means the system understands your entire research flow, not isolated searches. This is what Chrome’s AI Mode is aiming for, and it changes how we learn and