Each time we release a model, we run the same test: give it code that trains a small AI model, ask the new model to speed it up. It takes a skilled human 4-8 hours to reach 4x faster. In May 2024, Claude Opus 4 averaged a ~3x speedup. This April, Mythos Preview achieved ~52x.
TOOLS
-
New memory system enables review and control of ChatGPT’s context
By
–
With the new memory system, you can review and steer what ChatGPT remembers through a memory summary, with more visibility and control over how context is used. pic.twitter.com/kXMAds0g3q
— OpenAI (@OpenAI) 4 juin 2026With the new memory system, you can review and steer what ChatGPT remembers through a memory summary, with more visibility and control over how context is used.
-
Flows Agent lets you iterate and modify the pipeline dynamically
By
–
Flows Agent lets you iterate through conversation. Tell it to try a warmer voice, swap the background, or generate a version in Spanish. The agent modifies the pipeline and re-runs without rebuilding from scratch.
-
ElevenCreative Flows links 50+ models with voice, music, SFX
By
–
ElevenCreative Flows connects 50+ image and video models with voice, music, and SFX on one canvas. Creators and marketers use it to chain modalities together into complete pipelines and test creative variants across products, languages, and formats.
-
ElevenLabs introduces Flows Agent for automated creative workflow generation
By
–
Introducing Flows Agent in ElevenCreative.
— ElevenLabs (@ElevenLabs) 4 juin 2026
Describe what you want to create and the agent builds your entire workflow – selecting models, creating nodes, wiring connections, and running the generations. pic.twitter.com/m5Dyr77J15Introducing Flows Agent in ElevenCreative. Describe what you want to create and the agent builds your entire workflow – selecting models, creating nodes, wiring connections, and running the generations.
-
Using LLMs to substantiate private allegations with public evidence
By
–
I have a talk at Manifest about ParaLLM Construction coming up, about recent reporting I did which used LLMs to find public evidence substantiating (and letting me use) private allegations.
-

NVIDIA Nemotron 3 Ultra: open 550B hybrid Mamba-Attention MoE
By
–


1/ NVIDIA shipped Nemotron 3 Ultra today, a fully open 550B model with 55B active params, with the weights, training data, and complete recipe all released openly. That alone is rare at this scale. The headline however actually is speed. Ultra is a hybrid Mamba-Attention MoE, an
-

How a root agents.md ensures @agents .md is always created
By
–
in claude md just @agents
.md always i have a root agents md that tells agents when setting up new directories or repos to always do that. -

LangChain Labs study with Harvey on verifier efficiency benchmarking
By
–
In our LangChain Labs study with @Harvey
, we looked at how to measure efficiency across verifier designs. We benchmarked 5 setups against Sonnet per-criterion as the reference. -
Full study on efficient verifiers for legal agents
By
–
Read our full study: https://
langchain.com/blog/designing
-efficient-verifiers-for-legal-agents
…?