Y FUAH! Independientemente de lo costoso del modelo y tal, esta es una evidencia clara de que esto no para, y sobre todo en programación. Habiendo asumido un ritmo rápido pero progresivo con cada nuevo modelo, sorprende ver un salto tan bestia de golpe. Curvas vienen
LLMS
-
Anthropic Releases Mythos Preview Model Card Documentation
By
–
Aquí un link al model card del modelo Mythos Preview https://
www-cdn.anthropic.com/53566bf5440a10
affd749724787c8913a2ae0841.pdf
… -

Anthropic Mythos Preview: Major Breakthrough in Programming and Reasoning
By
–
¡ANTHROPIC MYTHOS PREVIEW! Anthropic acaba de publicar lo que serían los primeros benchmarks de su "filtrado" gran próximo modelo, Mythos. La verdad es que en programación y razonamiento el salto es BESTIA!
-

Claude Mythos Breaks SWE-Bench Pro with 77.8% Score
By
–

🚨 ANTHROPIC JUST BROKE SWE-BENCH PRO WITH CLAUDE MYTHOS 🚨 Anthropic just dropped the numbers for their unreleased "Claude Mythos Preview" and the coding leap is almost incomprehensible. This model is so powerful at finding exploits that they are keeping it strictly locked down for critical infrastructure partners. Anthropic explicitly stated: "We’ve used Claude Mythos to demonstrate thousands of zero day vulnerabilities." Look at the absolute destruction of these benchmarks compared to Opus 4.6: • SWE-Bench Pro: 77.8% (Destroying Opus 4.6 at 53.4%) • Terminal-Bench 2.0: 82.0% (Up from 65.4%) • SWE-Bench Verified: 93.9% • SWE-Bench Multimodal: 59.0% (More than double Opus 4.6's 27.1%) • Humanity's Last Exam (with tools): 64.7% (Up from 53.1%) • GPQA Diamond: 94.6% A nearly 25-point jump in SWE-Bench Pro in a single generation. And we’re in *checks notes* April..
→ View original post on X — @scobleizer, 2026-04-07 18:20 UTC
-

Claude Mythos Preview shows massive performance jump over Opus 4.6
By
–


This is beyond insanity. That jump is nuts. Opus 4.6 was released a few months ago. Look at that jump!! I am shocked Alex Albert (@alexalbert__) We released Claude Opus 4.6 just two months ago. Today we're sharing some info on our new model, Claude Mythos Preview. — https://nitter.net/alexalbert__/status/2041579938537775160#m
→ View original post on X — @kimmonismus, 2026-04-07 18:20 UTC
-
Anthropic Launches Mythos Preview for Project Glasswing Partners
By
–
Official Anthropic post Alex Albert (@alexalbert__) Mythos Preview is currently available to our launch partners in Project Glasswing. Learn more about the model and the project here: anthropic.com/glasswing — https://nitter.net/alexalbert__/status/2041579950332113155#m
→ View original post on X — @kimmonismus, 2026-04-07 18:18 UTC
-

Anthropic warns of serious damage risks from upcoming LLMs
By
–

Anthropic is being serious: they are afraid their upcoming LLMs could do serious damage. No end in sight „not long before such capabilities proliferate“ Chubby♨️ (@kimmonismus) MYTHOS BENCHMARKS, OFFICIAL. HOLY MOLY Anthropic cooked!! — https://nitter.net/kimmonismus/status/2041580372048187449#m
→ View original post on X — @kimmonismus, 2026-04-07 18:17 UTC
-
Claude Mythos Preview System Card Now Available
By
–
The Claude Mythos Preview system card is available here: anthropic.com/claude-mythos-preview-system-card [Translated from EN to English]
→ View original post on X — @anthropicai, 2026-04-07 18:15 UTC
-

Claude Mythos: New SWE Model with Impressive Progress
By
–


Claude MYTHOS: SWE verified, 93.9%, about 13% jump compared to Opus 4.6 WTF insane Alex Albert (@alexalbert__) We released Claude Opus 4.6 just two months ago. Today we're sharing some info on our new model, Claude Mythos Preview. — https://nitter.net/alexalbert__/status/2041579938537775160#m [Translated from EN to English]
→ View original post on X — @kimmonismus, 2026-04-07 18:15 UTC
-

Myth Benchmarks Official Results Anthropic Performance
By
–
MYTHOS BENCHMARKS, OFFICIAL. HOLY MOLY Anthropic cooked!!
→ View original post on X — @kimmonismus, 2026-04-07 18:14 UTC