AI Dynamics

Global AI News Aggregator

About

STARTUPS

  • AI Agents Managing Startups: YC-Bench Tests Profitability and Survival
    AI Agents Managing Startups: YC-Bench Tests Profitability and Survival

    Can an AI agent run a startup for a year without going bankrupt? Turns out most can't. New benchmark from Collinear AI puts 12 models to the test. YC-Bench tasks agents with running a simulated startup over hundreds of turns: hiring employees, selecting contracts, and maintaining profitability in a partially observable environment with adversarial clients and compounding consequences. Only three models consistently surpass the $200K starting capital. Claude Opus 4.6 leads at $1.27M average final funds, followed by GLM-5 at $1.21M with 11x lower inference cost. Scratchpad usage, the sole mechanism for persisting information across context truncation, is the strongest predictor of success. Adversarial client detection accounts for 47% of bankruptcies. Long-horizon coherence, not raw intelligence, separates the winners from the bankrupt. Paper: arxiv.org/abs/2604.01212 Learn to build effective AI agents in our academy: academy.dair.ai/

    → View original post on X — @dair_ai, 2026-04-02 15:37 UTC

  • $6M ARR with AI and One Employee at Polsia

    $6M+ ARR. One human employee. Almost 5,000 companies running 24/7 — while he sleeps. That's what @Bencera built with @polsia. New episode of Exponential Scale is live. If you're building with AI, this one's going to hit different. 🎧 Listen: scalebrate.com/podcast/6m-arr-with-all-agents-and-1-employee-interview-with-ben-broca-founder-of-polsia [Translated from EN to English]

    → View original post on X — @rschmelzer, 2026-04-02 15:01 UTC

  • Krea AI Launches Skills Package for Agent Integration

    introducing Krea Skills. any agent can now use Krea with just one command. npx skills add krea-ai/skills

    → View original post on X — @krea_ai

  • Mistral AI Secures $830M Funding Round Amid US-Europe Competition

    “The US is building two Apollo programs a year. Europe is building excellent regulation.” That’s not my line—it’s from the co-founder of @MistralAI
    , Arthur Mensch (
    @arthurmensch
    ). On Monday, Mistral announced an impressive $830 million financing round to build a cutting-edge

    → View original post on X — @ninadschick

  • Sakana releases interesting updates with Marlin beta launch

    Sakana keeps shipping interesting stuff. Will definitely check out Marlin beta!

    → View original post on X — @whats_ai

  • Sakana Marlin: AI Assistant Conducting 8-Hour Deep Research
    Sakana Marlin: AI Assistant Conducting 8-Hour Deep Research

    We are recruiting beta testers for Sakana Marlin🎣 This is a highly capable assistant with Deep Research that investigates over 8 hours! Behind this product lies AB-MCTS, the result of pure research projects! This is a product unique to Sakana AI, where research achievements translate into actual products👏 — Sakana AI (@SakanaAILabs) 🐟Ultra Deep Research Assistant "Sakana Marlin" – Now Recruiting Beta Testers🐟 Sakana AI has developed "Sakana Marlin," our first commercial product – an autonomous AI research assistant for business powered by our proprietary agent technology. sakana.ai/marlin-beta Sakana Marlin is an autonomous research assistant based on our unique long-term reasoning technology, capable of completing advanced business research. Key Features
    ・When given a topic, it autonomously conducts research for nearly 8 hours
    ・Automatically generates detailed research documents and summary slides
    ・Designed to replicate professional strategic research that teams of multiple people would conduct over weeks We conceived this solution to leverage AI's full potential to enable sound judgments in complex social situations. This technology fuses insights from "AI Scientist" – the automation of scientific discovery recently published in Nature magazine – with "AB-MCTS" which enables strategic exploration. This achieves "efficient reasoning scaling" where output quality improves the longer the AI thinks. We are conducting a closed beta test targeting those working daily on advanced research: strategy and business planning departments at financial institutions and corporations, consulting firms, think tanks, and similar organizations (free during the beta period). We will continuously improve based on your feedback. ▼ Apply for Closed Beta Tester Status Here
    forms.gle/MYHGP1wi2q4PHYPA7 [Translated from EN to English]

    → View original post on X — @sakanaailabs, 2026-04-02 08:48 UTC

  • Sakana AI Launches Marlin, AI Research Assistant for Business Beta Testing
    Sakana AI Launches Marlin, AI Research Assistant for Business Beta Testing

    New Ultra Deep Research Assistant Marlin @SakanaAILabs 🐠 Pushing the limits of test-time scaling for automating business-oriented research. It builds on top of AB-MCTS and The AI Scientist! Very excited to see agents scale to real-world applications and long-running workloads. Sign up for beta testing. — 🐟Ultra Deep Research Assistant "Sakana Marlin" – Beta Testers Wanted 🐟 Sakana AI has developed "Sakana Marlin," an AI research assistant for business use powered by proprietary agent technology, as our first commercial product. sakana.ai/marlin-beta Sakana Marlin is an autonomous research assistant based on our proprietary long-horizon reasoning technology for conducting advanced business research. Key Features
    ・ Conducts autonomous research for nearly 8 hours given a theme
    ・ Automatically generates detailed research documents and summary slides
    ・ Designed for professional strategic research that typically takes teams of multiple people several weeks We conceived this solution to maximize AI's potential for making high-quality decisions amid complex global circumstances. This technology fuses insights from "AI Scientist" (automated scientific discovery published in Nature magazine) with "AB-MCTS" (strategic exploration). It realizes "efficient inference scaling" where output quality improves with more reasoning time. Closed Beta Testing
    Targeted at professionals in financial institutions, corporations' strategic planning/business development divisions, consulting firms, think tanks, and similar organizations engaged in advanced research daily (free during the beta period). We will continuously improve based on your feedback. ▼ Apply for Closed Beta Testing
    forms.gle/MYHGP1wi2q4PHYPA7 [Translated from EN to English]

    → View original post on X — @sakanaailabs, 2026-04-02 08:38 UTC

  • Sakana AI Announces First Commercial Product Sakana Marlin Beta Test
    Sakana AI Announces First Commercial Product Sakana Marlin Beta Test

    We are proud to announce Sakana AI's first commercial product! 🐟 "Sakana Marlin" is an autonomous "Ultra Deep Research" assistant that functions as your Virtual CSO. By simply providing a single theme, it leverages the autonomous workflow technology of the "AI Scientist" published in Nature magazine last week and our proprietary reasoning scaling to autonomously think deeply and explore for up to 8 hours. Instead of mere surface-level summaries, it completes in one day the kind of advanced strategic research that professional research teams spend weeks on, generating detailed reports and presentation slides. Please read our blog article and consider applying to become a closed beta tester! 👇 — 🐟Ultra Deep Research Assistant "Sakana Marlin" – Beta Tester Recruitment 🐟 Sakana AI has developed "Sakana Marlin," an AI research assistant for business use powered by proprietary agent technology, as our first commercial product. sakana.ai/marlin-beta Sakana Marlin is an autonomous research assistant based on our unique long-term reasoning technology that accomplishes advanced business research. Key Features
    ・ Given a theme, it autonomously conducts research for nearly 8 hours
    ・ Automatically generates detailed investigation documents and summary slides
    ・ Designed to replicate professional strategic research that multiple teams spend weeks on Conceived as a solution that maximizes AI's potential for making high-quality decisions in complex social circumstances. This technology combines insights from "AI Scientist," the automation of scientific discovery published in Nature magazine, with "AB-MCTS" enabling strategic exploration. It realizes "efficient reasoning scaling" where output quality improves the longer the AI thinks. Closed Beta Test Launch
    Target participants include those engaged in advanced research on a daily basis, such as strategic planning/business planning divisions at financial institutions and corporations, consulting firms, and think tanks (free during the beta period). We will continue to improve based on your feedback. ▼ Apply for Closed Beta Tester
    forms.gle/MYHGP1wi2q4PHYPA7 [Translated from EN to English]

    → View original post on X — @sakanaailabs, 2026-04-02 08:28 UTC

  • Anduril EagleEye Helmet Lets Soldiers See Through Walls

    American soldiers can now see through walls. This isn't a video game. It's real technology from Anduril Industries. How it works: the EagleEye helmet fuses in real time data from Ghost-X drones and sensors deployed on the ground, then projects it all directly into the soldier's

    → View original post on X — @vision_ia

  • Atlas 1: Willow’s Speech-to-Text Model Tops Transcription Leaderboard

    Meet Atlas 1 — Willow's new frontier speech-to-text model just dropped, and it's already rewriting the leaderboard. It outperforms ElevenLabs, Deepgram, and OpenAI on transcription accuracy. Not by a little. By a wide margin. Built on the first scalable, human-powered

    → View original post on X — @futurepedia_io