Top stories in AI today: – xAI releases Grok 4 following 3’s crashout
– Perplexity’s Comet for AI-first web
– Turn messy image filenames into descriptive ones
– OpenAI poaches top engineers from rivals
– 4 new AI tools & 4 job opportunities Read more: https://
therundown.ai/p/xais-grok-4-
arrives
…
LLMS
-

Grok 4 launches as OpenAI recruits top AI talent
By
–
-
AGI Benchmarks Double Score in Unprecedented Leap
By
–
Un x2 sur les benchmark qui testent l'AGI, c'était jamais arrivé. Là c'est beaucoup d'un coup
-

Elon Musk Presents Grok 4 Leading a Beverage Company
By
–
Si vous ne l’avez pas encore vue, voici l’annonce officielle, ENTIÈREMENT TRADUITE sur VISION IA. Youtube : https://
youtu.be/R1BYiOv-Rgg C’est Elon Musk lui-même qui y présente Grok 4 et il va loin : il l’a carrément mis à la tête d’une entreprise de boissons… +5 000 € en un -
AI Intelligence Costs Decrease While Computational Demands Rise
By
–
De hecho lo es. El costo de la inteligencia a la que accedíamos a determinado precio hace un año es (mucho) más barata ahora. Lo que no impide que los modelos también puedan volverse mucho más potentes con usos más intensivos de computación, y por tanto más caros 🙂
-
Grok 4 evaluation results on ARC-AGI benchmark
By
–
Info chula sobre la evaluación de Grok 4 en ARC-AGI
-
Huggingface explains stateless direct response MCP server choice
By
–
If you're developing MCP servers, you should give a read to how the @huggingface team built the Hub MCP, they explain why they chose a Stateless + Direct Response server over other options!
-

Grok 4 Best Version Shows AGI Signs but Costs $300/Month
By
–
Pourquoi ELon, pourquoi La meilleure version de grok 4 qui montre des signes d'AGI, coute 300$ par mois… Après open ai et gemini, on dirai que c'est devenu la norme dans l'industrie
-

Tencent’s ArtifactsBench: Automated Evaluation for LLM Visual Artifacts
By
–
Tencent's Hunyuan team introduced ArtifactsBench, an automated evaluation pipeline for LLM-generated visual artifacts It assesses models on 1,825 diverse tasks with MLLM-as-Judge evaluating visual artifacts, achieving 94.4% ranking consistency with human experts
-
Ai2 Launches FlexOlmo: Decentralized Language Model Co-Development
By
–
Ai2 dropped FlexOlmo, a new paradigm for the co-development of language models
— The Rundown AI (@TheRundownAI) 10 juillet 2025
It allows each data owner to locally branch from a shared public model, add an expert trained on their data locally, and contribute this expert back to the shared modelpic.twitter.com/xiaaBKKOTSAi2 dropped FlexOlmo, a new paradigm for the co-development of language models It allows each data owner to locally branch from a shared public model, add an expert trained on their data locally, and contribute this expert back to the shared model
-

Grok 4 Benchmarks Outperform O3, Gemini 2.5 Pro, Claude Opus 4
By
–

XAI GROK 4 BENCHMARKS: OpenAI O3 is smoked
Gemini 2.5 Pro is smoked
Claude Opus 4 is smoked IT'S OVER, GROK 4 HAS WON.