Returning to the System Card, here is one of my favorite graphs, and it shows how the Mythos category models break the expected line of progress. As always with AI, we reach higher levels ahead of time.
MACHINE LEARNING
-

Cost frontiers: more expensive and superior model, Fable surpasses Opus 4.8
By
–
The cost frontiers, as expected, show an upward and rightward shift, meaning we have a more expensive and notably superior model (the low of Fable remains well above the x-high of Opus 4.8)
-

Benchmark tables: the model dominates in everything evaluated
By
–

Here are the complete benchmark tables. The model basically dominates in everything evaluated in these tables (Humanity Last Exam, CriPT, ArxivMath, HealthBench, etc.)
-
Language Swap by Pika MCP lets you speak any language in videos
By
–
Take your content global, already!
— Pika (@pika_labs) 9 juin 2026
The Language Swap skill via Pika MCP swaps the language you’re speaking in any video, making it look and sound like you speak fluent…anything. pic.twitter.com/kApd9yEbvXTake your content global, already! The Language Swap skill via Pika MCP swaps the language you’re speaking in any video, making it look and sound like you speak fluent…anything.
-
SambaNova shows disaggregated inference with up to 2x speed at Computex
By
–
Same prompt. Same model. Two stacks.
— SambaNova (@SambaNovaAI) 9 juin 2026
At #Computex, we demonstrated disaggregated inference live: GPUs handling prefill, SambaNova RDUs handling decode, and CPUs orchestrating agent execution.
The result? Up to 2X the speed of B200-only configurations 🦾 pic.twitter.com/YYP8o6WYrKSame prompt. Same model. Two stacks. At #Computex, we demonstrated disaggregated inference live: GPUs handling prefill, SambaNova RDUs handling decode, and CPUs orchestrating agent execution. The result? Up to 2X the speed of B200-only configurations
-

Build Reliable GenAI Applications with AI Evals, Observability, and Testing
By
–
Workshop hosted by @PacktPublishing @PacktDataML — "Build Reliable GenAI Applications with AI Evals, Observability, and Testing" 𝗥𝗲𝗴𝗶𝘀𝘁𝗲𝗿 𝗵𝗲𝗿𝗲 with my discount code 'KIRK40' (𝟰𝟬% 𝗢𝗙𝗙 already applied): https://
eventbrite.co.uk/e/build-reliab
le-genai-applications-with-ai-evals-observability-testing-tickets-1987301251540?aff=KirkB&discount=KIRK40
… 𝗟𝗲𝗮𝗿𝗻:
How production AI -

Joke-telling AI progress plateaued since 2022 PaLM
By
–
the new Fable still can't tell a joke i think jokery evals plateaued with Google's PaLM models in 2022, no one has pushed SOTA since then maybe another 10 trillion parameters will do the trick!
-

Cost of getting AI wrong: experiment first for production systems
By
–
Building AI Systems That Hold Up in Production — The Cost of Getting It Wrong: https://
odbms.org/blog/2026/06/t
he-cost-of-getting-it-wrong-ivan-santa-maria-filho-on-building-ai-systems-that-hold-up-in-production/
… via @odbmsorg Quote from article: "the most important lever is to experiment first, find exactly how AI will be used and whether it is the most cost effective way to solve -

Tackling Anthropic’s 319-page System Card
By
–
So far, the minimum we need to know. Now it's time to tackle the System Card of only 319 pages 🙂 Link https://www-cdn.anthropic.com/d00db56fa754a1b115b6dd7cb2e3c342ee809620.pdf …
-
Fable 5 available in Claude Code and Cowork, best coding model
By
–
Fable 5 is now available in Claude Code and Cowork
— Boris Cherny (@bcherny) 9 juin 2026
Fable is the best model I have used for coding, by a wide margin. It is a big step up, enabling less prompts and steers, more efficient token use, better code quality, better tool use, more intelligent self-verification, longer… https://t.co/RmVfZh39HtFable 5 is now available in Claude Code and Cowork. Fable is the best model I have used for coding, by far. It is a big step forward, allowing fewer instructions and guidance, more efficient token usage, better quality of
