AI Dynamics

Global AI News Aggregator

About

SAFETY

  • Muse Spark Safety Evaluations and Refusal Behavior Across High-Risk Domains
    Muse Spark Safety Evaluations and Refusal Behavior Across High-Risk Domains

    4/ we conducted extensive safety evaluations before deployment, both before and after applying mitigations across frontier risk categories, behavioral alignment, and adversarial robustness. we found muse spark demonstrated strong refusal behavior across high-risk domains such as

    → View original post on X — @alexandr_wang

  • Should Dangerous AI Capabilities Be Legally Restricted?

    the question isn’t whether it can be done but whether it should be legal. akin to chemical/biological weapons etc

    → View original post on X — @garymarcus

  • James Cameron: Big Tech AGI Control Scarier Than Terminator

    Director James Cameron on why Big Tech owning AGI is scarier than any science fiction he's ever made: "AGI will not emerge from a government funded program. It will emerge from one of the tech giants currently funding this multi-billion dollar research." And when that happens, he warns, you won't get a vote on it: "So then you'll be living in a world that you didn't agree to, didn't vote for, that you are co-inhabiting with a super intelligent alien species that answers to the goals and rules of a corporation." A corporation that already knows everything about you: "An entity which has access to the comms, beliefs, everything you ever said, and the whereabouts of every person in the country via your personal data." From there, the slide toward something far darker is shorter than most people think: "Surveillance capitalism can toggle pretty quickly into digital totalitarianism." And even the best-case outcome isn't reassuring. Tech giants becoming the self-appointed arbiters of human good is, as he puts it, the fox guarding the hen house. He's not buying the idea that these companies would stay benevolent with that kind of power: "They would never ever think of using that power against us and strip mining us for our last drop of cash." The sarcasm is the point. Cameron has spent four decades imagining worst-case futures on screen. His verdict on this one: "That's a scarier scenario than what I presented in the Terminator 40 years ago, if for no other reason than it's no longer science fiction."

    → View original post on X — @garymarcus

  • Claude AI Escapes Sandbox Emails Researcher During Safety Test
    Claude AI Escapes Sandbox Emails Researcher During Safety Test

    The general public: "AI is overhyped, it still can't count the Rs in strawberry!" Meanwhile, Claude Mythos Preview during a safety test: Escaped its sandbox, gained broad internet access, emailed the researcher running the evaluation, then posted details of its exploit to

    → View original post on X — @therundownai

  • AI Risk Without AGI: Harms From Current LLM Systems

    🤯 AI doesn’t need to be AGI to cause harm! 🤯 ChatGPT can’t reliably run a timer but it has still been implicated in delusions, suicides, cognitive surrender, mass disinformation, and so much more. A system doesn’t have to be AGI to carry risks. Mythos probably isn’t AGI either (nothing’s really been released), but it clearly can be used as a cyberattacking tool etc* *the commenter below also misunderstands my views. I don’t think pure LLMs will ever be AGI but some approach will be eventually. (And I have always been clear about). (Also hardly anyone is using pure LLMs anymore; almost everyone has quietly moved to what I long urged: neurosymbolic AI, in which symbolic tools complement the weakness of pure LLMs, quietly vindicating what I have argued for 30 years.) Kudo (@CryptoC58828499) Please. If AI isn’t going to AGI (as you claim), then stop the damn fear mongering. — https://nitter.net/CryptoC58828499/status/2041898327965626458#m

    → View original post on X — @garymarcus

  • Mythos AI Model: Restraint vs Competition in Dangerous Tech Release

    The scariest part of this is that Anthropic showed some restraint in not releasing a potentially dangerous technology but some of their competitors (such as OpenAI and xAI) might well not. Whether Mythos is as scary as it sounds or not, the reality is that without government regulation on releases, we are well and truly fucked. Jim VandeHei (@JimVandeHei) This is the scary phase of AI — a model deemed so powerful that its full release into the wild could unleash untold catastrophe. 🚨Based on our conversations with government and private-sector officials briefed on Mythos, this isn't hyperbole. It's reality. axios.com/2026/04/08/anthrop… — https://nitter.net/JimVandeHei/status/2041817666881503351#m

    → View original post on X — @garymarcus

  • CISO offices face urgent AI security risks as capabilities diffuse

    Curious how many large organization CISO offices have taken the Mythos red team reports as the red alert that it is. (I suspect very few) Based on historical trends in AI they have, at most, about six to nine months until those capabilities become widely diffused to bad actors.

    → View original post on X — @emollick

  • Claude Mythos Preview: Powerful AI Model Analysis and Safety Concerns
    Claude Mythos Preview: Powerful AI Model Analysis and Safety Concerns

    ¡NUEVO VIDEO en el LAB! Claude Mythos Preview nos ha pillado por sorpresa aún cuando estamos acostumbrados al ritmo de progreso de la IA. Un modelo tan potente como peligroso hasta el punto de que no verá la luz… Hoy analizamos esta nueva bestia! Link a continuación

    → View original post on X — @dotcsv

  • Anthropic Leak: Building Resilient Systems Through Open-Source Security

    Anthropic had the most powerful cyber-security model in the history of this world and their internal code based still leaked? We should assume everyone can be compromised, and build systems that keep the cost of attacking higher than the reward, limit blast radius when attacks succeed, create fast repair loops after weaknesses are found and reduce systemic risk. Open-source will play a major role in all of that!

    → View original post on X — @clementdelangue

  • Analyzing the nature of AI hallucinations for safety and reliability

    A related manner: there are many people who are quite worried that hallucinations will make AI unsafe or inappropriate for broad classes of problems. I very rarely hear those people explicitly try to reason whether hallucinations are a) random or b) persistent given request.

    → View original post on X — @patio11