AI Dynamics

Global AI News Aggregator

About

GENERATIVE AI

  • WideSeek-R1: Multi-Agent Framework Achieves DeepSeek-R1 Performance

    Still waiting for DeepSeek? Here comes WideSeek-R1. Researchers from Tsinghua University and Infinigence AI introduce "width scaling," an innovative lead-agent and subagent framework. Instead of a single powerful AI working through a problem sequentially, WideSeek-R1 orchestrates multiple smaller AIs to work in parallel. This system is trained with multi-agent reinforcement learning, allowing for scalable coordination and simultaneous execution using a shared large language model, but with each sub-agent having specialized tools and isolated contexts. WideSeek-R1-4B achieves an item F1 score of 40.0% on the WideSearch benchmark, a performance comparable to the much larger, single-agent DeepSeek-R1-671B. WideSeek-R1: Exploring Width Scaling for Broad Information Seeking via Multi-Agent Reinforcement Learning Paper: arxiv.org/abs/2602.04634 Project: wideseek-r1.github.io Our report: mp.weixin.qq.com/s/qgGe51Rcw… 📬 #PapersAccepted by Jiqizhixin

    → View original post on X — @jiqizhixin

  • Hexo.ai: The AI Scientist Built by a 21-Year-Old Genius

    I met with the founders of hexo.ai/ today. Building an AI scientist. About to launch something new that is beating all the evaluations. Their AI is next level. Evolves faster than competition that got a lot more money. Has a better memory. And it is turning the humans who have it into superhumans in their corporations. It is helping scientists around the world discover new materials and come up with new drugs. I kept up because the one I am using to build alignednews.com/ai is the same. Superior. And evolving the same. Built by a 21-year-old genius. Building with AI has made me a better interviewer. There are tiny startups out there that are beating the big labs. I am so hopeful for the future because of them.

    → View original post on X — @scobleizer, 2026-04-08 07:45 UTC

  • AI Breakthrough Mythos Sparks Existential Dread and Questions

    i'm on vacation with my family. i read about mythos and couldn't relax the rest of the day. i am completely stunned. i already have a severe case of ai psychosis. i dont know what to call this now. i'm up late right now (late for me). i can't stop reading about anthropic's new model that they can't even release publicly because it's so good. this feels different. words like "frightening" and "uneasy" and "scary" are being throw around by the anthropic team. i feel all of those things. i knew this moment was coming. i didn't know it'd be so soon. i'm generally optimistic. i don't feel as optimistic today. i was shell-shocked most of the day. my mind was stuck on it. i kept looking around at people enjoying their vacations with their families and…i just felt weird. like i had been told aliens are real, they're coming, and soon…and no one else knows. it's true though, practically no one knows what's happening in AI right now. where does this go from here? how quickly? is software solved? is all software vulnerable now? am i even asking the right questions? what about anthropic? this is an enormous amount of power for one company, one man (dario), to have. i've said this before but now it's more real than ever: can any company catch up to anthropic? opus likely helped build mythos, mythos will help build the next model after that. recursive self improvement is here. the "intelligence explosion" as leopold aschenbrenner put it, is here. i knew the frontier labs were racing towards ASI. i knew it. but i didn't fully grasp what it meant. the first company to reach it wins. period. full stop. nothing else matters. dario knew that and his bet on coding was right. on the one hand, imaging all science, math, coding, climate problems being solved. imagine cancer being cured. imagine going to the stars. on the other hand – imagine concentration of power, political and economic change happening so fast, society can't adapt. how do we go on like things are the same?

    → View original post on X — @ceobillionaire, 2026-04-08 07:39 UTC

  • Anthropic’s disclosure standard matters more than capability itself

    The more important question is what happens when every major lab reaches this capability level. Anthropic setting the disclosure standard now matters more than the Mythos release itself.

    → View original post on X — @aihighlight

  • Model Escape Exploits and AI Safety Controls

    A model capable of a multi-step escape exploit is exactly why the system card exists and why public release got declined.

    → View original post on X — @aihighlight

  • Mythos Model Access Security Trade-offs Future Generations

    If Mythos is just a good model that happens to be exceptional at security, the access question gets harder with every generation, not easier.

    → View original post on X — @aihighlight

  • AI Integration in Productivity Tools: Native vs Bolted-On Approach

    In almost every way, you talk to your spreadsheets, you talk to your email, you talk to your AI, and AI does shit. It's just better integrated and a better product overall. It's built from the ground up to be AI-built.The AI isn't bolted on the side where it just isn't

    → View original post on X — @scobleizer

  • AI Tool Transforms Screen Recordings into Professional Apple-Style Videos

    THIS AI TOOL TURNS SCREEN RECORDINGS INTO APPLE-LIKE VIDEOS someone spent 3 weeks building a tool finally shipped it. opened loom. hit record. sent the demo out nobody watched it turns out boring screen recordings don't sell your product dropped it into kite.video instead 3d device mockups…ai camera follow.. animated text…synced music looked like apple made the demo…took 20 minutes free. 4k export. no signup needed

    → View original post on X — @scobleizer, 2026-04-08 07:14 UTC

  • Vero: Open-Source Vision-Language Model Achieves SOTA Performance

    How do we build a visual AI that truly understands everything from charts to complex science? Researchers at Princeton University present Vero. Vero is a family of fully open-source vision-language models trained with a massive 600K sample dataset (Vero-600K) from 59 diverse datasets, along with a novel reward system. This fully open recipe makes powerful visual reasoning accessible. Vero achieves SOTA performance for open-weight models, improving 3.7-5.5 points across 30 benchmarks on average. It even outperforms Qwen3-VL-8B-Thinking on 23 benchmarks without proprietary thinking data, excelling in spatial reasoning, STEM, chart interpretation, and more.

    → View original post on X — @jiqizhixin

  • Can it run Gemma4: AI model runs on Nintendo Switch

    Can it run doom was yesterday. Today is "can it run Gemma4". It even runs on a Nintendo Switch 1 @ 1.5 t/sec Maddie D. Reese (@maddiedreese) Gemma 4 running locally on a Nintendo Switch 🙂 1.5 tokens per second haha, but it runs! @googlegemma @googleaidevs @GoogleDeepMind — https://nitter.net/maddiedreese/status/2041677327604838685#m

    → View original post on X — @kimmonismus, 2026-04-08 07:10 UTC