Excited to release our technical report: “The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search” https://
pub.sakana.ai/ai-scientist-v
2/paper/
… The AI Scientist-v2 incorporates an “Agentic Tree Search” approach into the workflow, enabling deeper and more
LLMS
-

AI Scientist-v2: Agentic Tree Search for Automated Scientific Discovery
By
–
-

Creator Pauses Llama 4 Analysis Amid Conflicting Reports
By
–
Been a busy few months. Gonna check out for a bit. I made a whole video about Llama 4 and decided not to release it as more details come out around what was announced vs what results people are seeing. Will revisit it after the story’s finished unfolding. For now, gonna jump
-
Identifying AI Models Behind APIs and Switching Economics
By
–
How do you know which models are behind an API and why they wouldn’t switch them?
-

Actual LLM Agents Coming: Latest Developments
By
–
Actual LLM agents are coming | Vintage Data https://
buff.ly/xcDK3c4 #AI #MachineLearning #DeepLearning #LLMs #DataScience -

Scaling Laws of Synthetic Data for Language Models
By
–
Scaling Laws of Synthetic Data for Language Models Qin et al.: https://
arxiv.org/abs/2503.19551
v2
… #ArtificialIntelligence #DeepLearning #MachineLearning -
Chinese AI Models Challenge US Dominance in Global Competition
By
–
The AI competition is heating up, with rising Chinese models challenging the US lead, and narrowing performance gaps between top models, according to the #AIIndex2025 report. Read more via @Nature
: -

APIGen-MT: Agentic Pipeline for Multi-Turn Data Generation
By
–
APIGen-MT: Agentic Pipeline for Multi-Turn Data Generation via Simulated Agent-Human Interplay Prabhakar et al.: https://
arxiv.org/abs/2504.03601 #ArtificialIntelligence #DeepLearning #MachineLearning -

USAMO Solutions Analysis Model Improvements Comparison
By
–
I somehow missed this, wow! Any analysis from the solutions on USAMO of the improvements vs other models?
-

Groq Llama 4 Scout Achieves High Performance Across Context Lengths
By
–
Current verified price performance by @artificialanalysis, with consistent high performance across context lengths: https://
artificialanalysis.ai/models/llama-4
-scout/providers
…