This is simply amazing demonstration of Pika's real time video chat. Appreciate it!
MULTIMODAL AI
-

Qwen3.6 Plus Vision-Language Model Now Available on Poe
By
–
Qwen3.6 Plus is now live on Poe. Delivers advanced vision‑language performance from Qwen, with clear gains in code‑heavy workflows like agentic and front‑end coding, plus stronger multimodal understanding including improved OCR and precise object localization. Designed for
-
Point Tracks Enable Long-Range Animal Motion Forecasting in World Models
By
–
Rapid progress in world model space. This paper uses point tracks as representation, enabling long range forecasting of animal motion. Led by my PhD student @neerjathakkar from her GDM internship. Neerja Thakkar (@neerjathakkar) What’s the right representation for a world model? 3D, pixels, or something else? Excited to release our new paper “Forecasting Motion in the Wild” where we propose point tracks as tokens for generating complex non-rigid motion and behavior From @GoogleDeepmind @Berkeley_AI @TTIC_Connect — https://nitter.net/neerjathakkar/status/2039701926980260205#m
→ View original post on X — @berkeley_ai, 2026-04-02 22:00 UTC
-
Pika Launches PikaStream1.0: Real-Time Video Chat Skill for AI Agents
By
–
Conversations tend to go better with a face and a voice. That’s why we’re thrilled to release the beta version of the first video chat skill for ANY agent, powered by our new real-time model, PikaStream1.0.
— Pika (@pika_labs) 2 avril 2026
The skill preserves memory and personality, and enables real-time… pic.twitter.com/IgC4vcB0T7Conversations tend to go better with a face and a voice. That’s why we’re thrilled to release the beta version of the first video chat skill for ANY agent, powered by our new real-time model, PikaStream1.0. The skill preserves memory and personality, and enables real-time adaptability. And if you use it with your Pika AI Self, they’ll be able to execute agentic tasks during the call 💅
→ View original post on X — @paulroetzer, 2026-04-02 20:38 UTC
-

Microsoft Launches Copilot Advisors AI Debate Experiment
By
–

Microsoft released Copilot Advisors experiment in the US where multiple AI characters with live portraits will debate on a certain topic. Users will be able to select from various personalities and specify any topic they want. At the end, the debate will be summarized on a
-

AI Trust and Microsoft’s MAI-Image-2 Model Achievement
By
–
The most meaningful AI work doesn’t just advance intelligence, it earns trust. Shrijayan (@rshrijayan) Microsoft's AI Superintelligence team just released MAI-Image-2, a text-to-image model that landed at No. 5 on the Arena AI leaderboard — marking the strongest release yet for Mustafa Suleyman’s lab. — https://nitter.net/rshrijayan/status/2034987076144468125#m
-
AI Reading X Posts For You: User Reactions and Perspectives
By
–
What is your reaction to having AI read X for you?
-

AI Transforms Media: NotebookLM Acquisition Analysis
By
–
AI is changing media.
— Robert Scoble (@Scobleizer) 2 avril 2026
Deeply.
Pay attention.
Here's the video from NotebookLM about @tbpn's acquisition and my analysis. And a LOT more. https://t.co/J4gvjfC4Dz pic.twitter.com/sThcV641ETAI is changing media. Deeply. Pay attention. Here's the video from NotebookLM about @tbpn
's acquisition and my analysis. And a LOT more. -
Google Releases Gemma 4: Open Source AI Models
By
–
🔴 ¡GOOGLE LIBERA GEMMA 4!
— Carlos Santana (@DotCSV) 2 avril 2026
Cuatro versiones abiertas muy interesantes:
👉 31B Dense y 26B MoE: rendimiento equivalente a alternativas más grandes en tamaños muy accesibles!
👉 E4B y E2B: ligeros con procesamiento en tiempo real de texto, visión y audio!pic.twitter.com/aGhWF3zXFg¡GOOGLE LIBERA GEMMA 4! Cuatro versiones abiertas muy interesantes: 31B Dense y 26B MoE: rendimiento equivalente a alternativas más grandes en tamaños muy accesibles! E4B y E2B: ligeros con procesamiento en tiempo real de texto, visión y audio!