Meet Hark – the most ambitious AI lab you haven't heard of yet. Founded by serial entrepreneur Brett Adcock, Hark is building AI that's proactive, personalized, and speaks through voice, text, vision & memory. 45+ researchers from Apple, Meta, Google & Tesla iPhone's
@futurepedia_io
-
CapCut Video Studio Uses AI Agent for Full Video Creation Workflow
By
–
CapCut just reinvented video creation — no timeline required. Studio (powered by Dreamina Seedance 2.0) lets AI handle the whole workflow: AI agent brainstorms & builds your storyboard Generates scenes with consistent characters & faces Up to 15 seconds
-
ByteDance Releases DeerFlow 2.0 Open-Source SuperAgent on GitHub
By
–
ByteDance dropped an open-source SuperAgent — and it hit #1 on GitHub Trending. DeerFlow 2.0 can: Deep web research with cited sources Generate full reports with charts, images & video Run Python & bash in a secure sandbox Create slide decks & UI components
-
Genspark AI Workspace 3.0 Launches Voice Commands and AI Employee
By
–
Genspark wants to replace your keyboard with your voice. Realtime Voice is part of Genspark AI Workspace 3.0 – give voice commands, track tasks in real time, get instant results. Also launched: • Genspark Claw – their first AI employee • AI Slides, Docs & Sheets •
-
Freepik 3D Scenes Turns Product Photos Into 3D Models
By
–
Freepik just turned product photography into a 3D workflow – no studio required. 3D Scenes is Freepik's virtual scene builder, now fully integrated with their 3D Generator: Upload a product photo → get a production-ready 3D model (GLB) Drop it straight into a 3D
-
Google Integrates Gemini AI Into Marketing Platform at NewFront 2026
By
–
Google just brought Gemini into its Marketing Platform. Announced at NewFront 2026 – three new launches: • Ads Advisor – AI campaign setup, optimization & reporting in DV360 • Live Sports Biddable Suite – real-time ad bidding for live events
• Confidential Publisher -
Google Gemini 3.1 Flash Live Real-Time Audio Model Launched
By
–
Google's best real-time audio model just went live. Gemini 3.1 Flash Live is built for natural, low-latency voice conversations. Recognizes pitch, pace & acoustic nuance Filters background noise 90+ languages supported All audio watermarked with SynthID
-
Luma Launches Uni-1 Multimodal Image Model Beating Google and OpenAI
By
–
Luma just launched a multimodal image model that beats Google and OpenAI on benchmarks. Uni-1 is built on a decoder-only transformer trained on images, video, audio, language AND spatial reasoning. #1 in human preference Elo for Style & Editing 10–30% cheaper than
-
Anthropic Launches Claude Computer Use Feature
By
–
Claude can now USE your computer. Anthropic launched Computer Use – Claude clicks, scrolls, browses, and completes tasks for you. Opens files, browsers & dev tools Connects to Google Workspace & Slack Always asks permission before accessing new apps
-
Mistral Launches Voxtral Transcribe 2: Fast, Cheap, Open-Source
By
–
Mistral just upgraded Voxtral – and it's fast, cheap, and open. Voxtral Transcribe 2 comes in two flavors: • Mini – batch transcription, speaker diarization, word-level timestamps, 13 languages • Realtime – sub-200ms latency, open-source (Apache 2.0) Priced at