¡ANTHROPIC MYTHOS PREVIEW! Anthropic acaba de publicar lo que serían los primeros benchmarks de su "filtrado" gran próximo modelo, Mythos. La verdad es que en programación y razonamiento el salto es BESTIA!
AI
-

Claude Mythos Breaks SWE-Bench Pro with 77.8% Score
By
–

🚨 ANTHROPIC JUST BROKE SWE-BENCH PRO WITH CLAUDE MYTHOS 🚨 Anthropic just dropped the numbers for their unreleased "Claude Mythos Preview" and the coding leap is almost incomprehensible. This model is so powerful at finding exploits that they are keeping it strictly locked down for critical infrastructure partners. Anthropic explicitly stated: "We’ve used Claude Mythos to demonstrate thousands of zero day vulnerabilities." Look at the absolute destruction of these benchmarks compared to Opus 4.6: • SWE-Bench Pro: 77.8% (Destroying Opus 4.6 at 53.4%) • Terminal-Bench 2.0: 82.0% (Up from 65.4%) • SWE-Bench Verified: 93.9% • SWE-Bench Multimodal: 59.0% (More than double Opus 4.6's 27.1%) • Humanity's Last Exam (with tools): 64.7% (Up from 53.1%) • GPQA Diamond: 94.6% A nearly 25-point jump in SWE-Bench Pro in a single generation. And we’re in *checks notes* April..
→ View original post on X — @scobleizer, 2026-04-07 18:20 UTC
-

Claude Mythos Preview shows massive performance jump over Opus 4.6
By
–


This is beyond insanity. That jump is nuts. Opus 4.6 was released a few months ago. Look at that jump!! I am shocked Alex Albert (@alexalbert__) We released Claude Opus 4.6 just two months ago. Today we're sharing some info on our new model, Claude Mythos Preview. — https://nitter.net/alexalbert__/status/2041579938537775160#m
→ View original post on X — @kimmonismus, 2026-04-07 18:20 UTC
-
Autonomous Wheeled Bipedal Robot Navigates Complex Terrain Successfully
By
–
#Autonomous Wheeled Bipedal #Robot Masters Complex Terrain Navigation
— Ronald van Loon (@Ronald_vanLoon) 7 avril 2026
via @ZappyZappy7
#Innovation #Robotics #EmergingTech #Tech #Technology pic.twitter.com/erKhmiLUmq#Autonomous Wheeled Bipedal #Robot Masters Complex Terrain Navigation via @ZappyZappy7 #Innovation #Robotics #EmergingTech #Tech #Technology
→ View original post on X — @ronald_vanloon, 2026-04-07 18:19 UTC
-
Anthropic Launches Mythos Preview for Project Glasswing Partners
By
–
Official Anthropic post Alex Albert (@alexalbert__) Mythos Preview is currently available to our launch partners in Project Glasswing. Learn more about the model and the project here: anthropic.com/glasswing — https://nitter.net/alexalbert__/status/2041579950332113155#m
→ View original post on X — @kimmonismus, 2026-04-07 18:18 UTC
-

Anthropic warns of serious damage risks from upcoming LLMs
By
–

Anthropic is being serious: they are afraid their upcoming LLMs could do serious damage. No end in sight „not long before such capabilities proliferate“ Chubby♨️ (@kimmonismus) MYTHOS BENCHMARKS, OFFICIAL. HOLY MOLY Anthropic cooked!! — https://nitter.net/kimmonismus/status/2041580372048187449#m
→ View original post on X — @kimmonismus, 2026-04-07 18:17 UTC
-
AMD Said to Have Surpassed NVIDIA in AI Performance
By
–
Source: forbes.com/sites/karlfreund/2026/04/06/did-amd-just-beat-nvidia-in-ai-performance/ [Translated from EN to English]
→ View original post on X — @kimmonismus, 2026-04-07 18:17 UTC
-

MLPerf 6.0: NVIDIA Dominates AI Benchmarks, AMD Shows Progress
By
–

Did AMD Just Beat NVIDIA In AI Performance? No. And the article itself says so. I really don't like clickbaity headlines. They even state it at the very end, though the title suggest a bit otherwise: "NVIDIA ran every newly added benchmark and won every one of them. Only two of these were attempted by AMD, and for which Nvidia out-performed them by ~30 and ~50%. So, no, AMD did not beat Nvidia." MLPerf 6.0 results are out and the actual data tells a clear story: -NVIDIA won every new benchmark it entered -GB300 NVL72 delivered nearly 3x more throughput than 6 months ago, same hardware, better software -2.5M tokens/sec on DeepSeek R1 with 288 B300s AMD made real progress with the MI355X; getting within 10-30% on select single-node tests is no joke. And they deserve credits for that. But they skipped most new benchmarks and didn't compete on the hardest models. Imho / take: The real story isn't GPU vs GPU anymore. It's full-stack AI infrastructure: networking, software optimization, disaggregated serving. And tbh that's where NVIDIA keeps pulling ahead.
→ View original post on X — @kimmonismus, 2026-04-07 18:17 UTC
-
Claude Mythos Preview System Card Now Available
By
–
The Claude Mythos Preview system card is available here: anthropic.com/claude-mythos-preview-system-card [Translated from EN to English]
→ View original post on X — @anthropicai, 2026-04-07 18:15 UTC
-

Claude Mythos: New SWE Model with Impressive Progress
By
–


Claude MYTHOS: SWE verified, 93.9%, about 13% jump compared to Opus 4.6 WTF insane Alex Albert (@alexalbert__) We released Claude Opus 4.6 just two months ago. Today we're sharing some info on our new model, Claude Mythos Preview. — https://nitter.net/alexalbert__/status/2041579938537775160#m [Translated from EN to English]
→ View original post on X — @kimmonismus, 2026-04-07 18:15 UTC
