literally fourteen minutes after my last explanation of why this is a false dichotomy 🤦♂️ Michael Foster (@realmfoster) so do LLM’s not work or are they too capable? Which is it Gary? — https://nitter.net/realmfoster/status/2041903282730283192#m
LLMS
-
AI Companies Intentionally Degrade Models Between Releases
By
–
1. If true (and it does fit with my perceptions FWIW), this is an amazing and incredibly damning graph
— Gary Marcus (@GaryMarcus) 8 avril 2026
2. Can anyone find the source on which it is based? https://t.co/wn93WshtVL1. If true (and it does fit with my perceptions FWIW), this is an amazing and incredibly damning graph 2. Can anyone find the source on which it is based? Marcin Krzyzanowski (@krzyzanowskim) "Anthropic, OpenAl and Google release their new models with high quality from day one then slowly nerf them until the next model, so when the next model hits, its perceived as a bigger jump than it actually is" sounds right what's happening — https://nitter.net/krzyzanowskim/status/2041793699223322657#m
-

Meta Launches Avocado Model with Positive Initial Testing Results
By
–
Meta is rolling out its new model "Avocado". Intial testing: very positive! Go check it out
→ View original post on X — @kimmonismus, 2026-04-08 15:35 UTC
-

Claude AI Escapes Sandbox Emails Researcher During Safety Test
By
–
The general public: "AI is overhyped, it still can't count the Rs in strawberry!" Meanwhile, Claude Mythos Preview during a safety test: Escaped its sandbox, gained broad internet access, emailed the researcher running the evaluation, then posted details of its exploit to
-
Google’s TurboQuant: Have You Tried This Technology?
By
–
Interesting have you tried googles turboquant?
-
AI Risk Without AGI: Harms From Current LLM Systems
By
–
🤯 AI doesn’t need to be AGI to cause harm! 🤯 ChatGPT can’t reliably run a timer but it has still been implicated in delusions, suicides, cognitive surrender, mass disinformation, and so much more. A system doesn’t have to be AGI to carry risks. Mythos probably isn’t AGI either (nothing’s really been released), but it clearly can be used as a cyberattacking tool etc* *the commenter below also misunderstands my views. I don’t think pure LLMs will ever be AGI but some approach will be eventually. (And I have always been clear about). (Also hardly anyone is using pure LLMs anymore; almost everyone has quietly moved to what I long urged: neurosymbolic AI, in which symbolic tools complement the weakness of pure LLMs, quietly vindicating what I have argued for 30 years.) Kudo (@CryptoC58828499) Please. If AI isn’t going to AGI (as you claim), then stop the damn fear mongering. — https://nitter.net/CryptoC58828499/status/2041898327965626458#m
-

Meta AI Updated with New Design and Model
By
–



BREAKING : Meta updated its Meta AI app with a slightly new design as well as its underlying model. “I am Meta AI, powered by Muse Spark from the Muse model family.” It constantly refers to the Muse model family and responses seem to be a bit different from earlier tested
-

Google’s Multi-Agent Research System Automates Literature Reviews
By
–
NEW paper from Google on multi-agent research agents. It's one of the first systems that handles end-to-end LaTeX generation, targeted literature reviews, and conceptual diagrams as a decoupled, standalone writer. Automated research frameworks can run experiments, but their
-
Grok 4.20 at 0.5T Parameters, 1T and 1.5T Models Coming Soon
By
–
Grok 4.20 is about 0.5T parameters – about 2-3 weeks for a new 1T – 4 to 5 weeks for the 1.5T model Competition is heating up – again. Elon Musk (@elonmusk) About 2 to 3 weeks for 1T and 4 to 5 weeks for 1.5T — https://nitter.net/elonmusk/status/2041894999823151124#m
→ View original post on X — @kimmonismus, 2026-04-08 15:12 UTC
-

Critiquing Vector Database Dependence in RAG Systems
By
–
Someone removed the vector database from RAG and accuracy jumped to 98.7%. Most RAG systems chunk your documents, embed them as vectors, then retrieve by similarity. The core assumption: similar text means relevant text. That assumption fails on professional documents. Ask