To those in the replies who say "but opus 4.8 is weaker so without fallback, the score would be even higher": this is not necessarily true because of how any benchmark works – which is an average of queries – and what is called "the x.com/ClementDelangu…"
MACHINE LEARNING
-
Refutation of the argument that a weaker model increases the score
By
–
To the people in the replies who say "but opus 4.8 is weaker, so without fallback, the score would be even higher": this is not necessarily true because of how any benchmark works – which is an average of queries – and what is called "the x.com/ClementDelangu…"
-
Benchmarks: why a weaker model does not necessarily improve the score
By
–
To those in the replies who say 'but opus 4.8 is weaker so without fallback, the score would be even higher': this is not necessarily true because of how any benchmark works – which is an average of queries – and what is called 'the x.com/ClementDelangu…'
-

Asynchronous AI Cuts Energy by Orders of Magnitude while Learning Continuously
By
–
Asynchronous #AI cuts computing energy by orders of magnitude while learning continuously
by Daegan Miller @TechXplore_com Learn more: https://
bit.ly/4fuihmz #MachineLearning #ArtificialIntelligence #DL #ML -
Debate on the impact of fallback in benchmarks
By
–
To the people in the replies who say 'but opus 4.8 is weaker so without fallback, the score would be even higher': that is not necessarily true because of how any benchmark works – which is an average of queries – and what is called 'the x.com/ClementDelangu…'
-
Refutation on benchmarks and the lack of fallback
By
–
To the people in the replies who say "but opus 4.8 is weaker so without fallback, the score would be even higher": this is not necessarily true because of how any benchmark works – which is an average of queries – and what is called "the x.com/ClementDelangu…"
-
Refutation of the argument on fallback in benchmarks
By
–
To those in the replies who say 'but opus 4.8 is weaker so without fallback, the score would be even higher': this is not necessarily true because of how any benchmark works – which is an average of queries – and what is called 'the'
-
SambaNova congratulates MiniMax on M3 open-weight model launch
By
–
Congrats to our partners at @MiniMax_AI on the launch of MiniMax M3. Open-weight models continue to push the ecosystem forward, and we're excited to bring M3 to RDUs down the road. Looking forward to following what's built with it.
-
Databricks Genie expands to predictive analytics with TabPFN and Agent Bricks
By
–
Business users can already use Genie to ask descriptive questions in natural language. Now the same conversational workflow can support predictive analytics too.
— Databricks (@databricks) 12 juin 2026
By combining Databricks Genie, TabPFN, and Agent Bricks, teams can turn questions like “Which customers are likely to… pic.twitter.com/z71u2MoBw2Business users can already use Genie to ask descriptive questions in natural language. Now the same conversational workflow can support predictive analytics too. By combining Databricks Genie, TabPFN, and Agent Bricks, teams can turn questions like “Which customers are likely to
-
Progress in fine-grained 3D motion control for AI video
By
–
Fine-grained 3D motion control in AI video just got a little bit closer https://t.co/Uqi4lVJunR
— fofr (@fofrAI) 12 juin 2026Fine-grained 3D motion control in AI-generated video just got a little closer
