
Seems like Google is testing a new Gemma 3 75b model on lmarena under the "cutiepie-75" name.

By
–

Seems like Google is testing a new Gemma 3 75b model on lmarena under the "cutiepie-75" name.
By
–
Of those, ChatGPT's is the best. I think Gemini's Deep Research is even better than all of those though.
By
–
Yann LeCun ne croit pas que les LLM soient l’avenir de l’IA.
— VISION IA (@vision_ia) 15 mai 2025
Pour lui, c’est une impasse :
→ Pas de vraie compréhension du monde
→ Pas de mémoire persistante
→ Raisonnement trop simpliste
Il mise sur autre chose. Et si c’était lui qui voyait juste… cinq ans trop tôt ?… pic.twitter.com/ikGkZRj8fh
Yann LeCun ne croit pas que les LLM soient l’avenir de l’IA. Pour lui, c’est une impasse :
→ Pas de vraie compréhension du monde
→ Pas de mémoire persistante
→ Raisonnement trop simpliste Il mise sur autre chose. Et si c’était lui qui voyait juste… cinq ans trop tôt ?

By
–
Je ne comprends pas la stratégie d’Anthropic. On dirait simplement qu’ils ne veulent pas participer à la course vers le sommet. Ils ont un des meilleurs modèles dispo mais ils sont prêt à tout balancer à la poubelle. Traduction : Pourquoi Claude perd des utilisateurs ?
«

By
–
Turn any website into LLM-ready data with just a few clicks!
— Sumanth (@Sumanth_077) 15 mai 2025
Firecrawl just released Templates, a collection of ready-to-use playground setups, code snippets, and full repositories to scrape and structure web data for your projects.
Getting web data just got a lot easier. pic.twitter.com/rdelfEpYhl
Turn any website into LLM-ready data with just a few clicks! Firecrawl just released Templates, a collection of ready-to-use playground setups, code snippets, and full repositories to scrape and structure web data for your projects. Getting web data just got a lot easier.
By
–
Meta ne peut pas être sérieux avec l’IA. – Rendez-vous sur Meta AI pour tester Llama 4
– Ajoutez une image → le prompt l’ignore
– Essayez de vous connecter → échec, retour au logo Facebook Ce n’est plus un problème de Llama 4, de LLM vs modèles mondes, c’est un problème de

By
–
Excited to share insights from our latest guest post on Decoding ML by @iusztinpaul
! We walk through deploying DeepSeek‑R1 Distill on three infrastructure models—each with its own trade‑offs: 1. Google Cloud Platform (GCP) • Self‑managed VM with an NVIDIA L4 GPU
• Full
By
–
Can we trust LLMs in high-risk use cases? Join us for a webinar on DLBacktrace by @AryaXAI — a new model-agnostic XAI technique for deep learning & LLMs. 📅 Thursday, June 5, 2025 🕒 11:00 AM EDT | 4:00 PM BST | 8:30 PM IST Register now: aryaxai.com/events/inside-th…

By
–
Insights into DeepSeek-V3: Scaling Challenges and Reflections on Hardware for AI Architectures
Paper: https://
arxiv.org/pdf/2505.09343

By
–
New from DeepSeek! They just released a paper diving into the scaling challenges and hardware reflections behind DeepSeek-V3. As large language models (LLMs) scale, hardware becomes the bottleneck. DeepSeek-V3—trained on 2,048 NVIDIA H800 GPUs—offers a compelling case study