Mistral Small 4 is live on Poe! One model for reasoning, vision, and agentic coding. Good for things like solving multi-step math or research problems, extracting structured data from images or scanned documents, debugging and navigating a codebase, analyzing a chart and
LLMS
-
GPT-5.3-Codex-Spark: Three Real Workflows for Building
By
–
what can you build with gpt-5.3-codex-spark?@jxnlco from @OpenAI demos 3 real workflows — ones you can set up yourself inside the Codex app to help you spend less time on overhead and more time building.
— Cerebras (@cerebras) 18 mars 2026
00:09 – what is gpt-5.3-codex-spark?
00:25 – workflow 1: multi-agent… pic.twitter.com/ckpJREPJ3ewhat can you build with gpt-5.3-codex-spark? @jxnlco from @OpenAI demos 3 real workflows — ones you can set up yourself inside the Codex app to help you spend less time on overhead and more time building. 00:09 – what is gpt-5.3-codex-spark? 00:25 – workflow 1: multi-agent daily briefing from slack, drive & meets 01:06 – workflow 2: automated PR review 01:31 – workflow 3: real-time interactive coding 02:56 – what speed changes, and what's coming next
-

Microsoft Copilot to integrate deeper OneDrive context for improved responses
By
–
Copilot may also get a deeper OneDrive sync for Copilot Library. "Copilot will use context from your Microsoft OneDrive to improve answers."
-
Attention Head as a Differentiable Multiplexer
By
–
An attention head is a differentiable multiplexer.
-

LLM Sycophancy and Chatbot-Associated Delusions: A Mind-Blowing 37%
By
–
Holy crap. I knew about sycophancy. But the 37% number below blows my mind. This from an analysis of chat logs in people who experienced chatbot-associated delusions. In over a third of the messages to those users, the LLMs told the users they had (eg)
-

AI Agents Could Push College Graduate Unemployment Above 30%
By
–
#AIAgents could easily send college grad unemployment over 30%, ServiceNow CEO says
by @samantha_subin @cnbc Learn more: https://
bit.ly/3NfapcV #LLM #GenerativeAI #ArtificialIntelligence #MachineLearning -

ODSC AI East 2026: Production AI Engineering Practices and Techniques
By
–
At ODSC AI East 2026, the focus is learning the Engineering Practices behind Production AI systems: Agentic AI, RAG, LLMOps, LLMs, GenAI, Model Fine-Tuning, … Boston + Virtual on April 28–30 Register here: https://
hubs.li/Q04746MR0 via @_odsc -
Quantization Impact on Model Quality and Expert Reduction
By
–
How confident are you with respect to the output quality given the 2-bit quantization and reducing experts from 10 to 4? Did you have a mechanism for measuring that?
-

Running 397B MoE Model on M3 Mac with Efficient Weight Streaming
By
–
Dan says he's got Qwen 3.5 397B-A17B – a 209GB on disk MoE model – running on an M3 Mac at ~5.7 tokens per second using only 5.5 GB of active memory (!) by quantizing and then streaming weights from SSD (at ~17GB/s), since MoE models only use a small subset of their weights for
-

MiniMax-M2.7 model hits SWE-Pro SOTA at 56.22%
By
–

MiniMax-M2.7 is here. → matching Sonnet 4.6 as an agent
→ recovering live incidents in just 3 minutes
→ hitting SOTA in SWE-Pro (56.22%) It even edits your Office files
