We were ridiculed when we started using Ai avatars (“who will watch this, no emotion”) We were ridiculed on saying video editors would be a mainstream role (“bro they’re paid just 5-10k a month”) We were ridiculed on the first game trailer (“it’ll never get better”) And so
MULTIMODAL AI
-

Nano Banana Pro: Image Generation and AI Scam Trends 2026
By
–
Nano Banana Pro remains my favorite image generation model. Undoubtedly, 2026 will see more AI scams. And more revenue-generating AI influencers. Prompt: This woman is sitting on a white beanbag and laughing. She’s wearing a light pink cap, a black high necked crop top, and
-

Meta Open-Sources PE-AV for Advanced Audio Separation Technology
By
–
We’re open-sourcing Perception Encoder Audiovisual (PE-AV), the technical engine that helps drive SAM Audio’s state-of-the-art audio separation. Built on our Perception Encoder model from earlier this year, PE-AV integrates audio with visual perception, achieving
-

xAI to launch Imagine API and Playground for Grok models
By
–
BREAKING : xAI is working on Imagine API and API Playground! – Imagine API will expose Grok image and video models to developers. – Playground will let developers play with Grok models and tweak different parameters to see how it performs. Grok AI Studio
-
Luma Labs Launches Ray 3 Modify for Video Generation
By
–
BREAKING 🚨: Luma Labs is launching Ray 3 Modify, a new video generation model that allows users to modify existing videos by providing character reference images.
— 🚨 AI News | TestingCatalog (@testingcatalog) 18 décembre 2025
You can inject yourself everywhere now 👀 https://t.co/83eR5llUBu pic.twitter.com/i8Bp25zNSlBREAKING : Luma Labs is launching Ray 3 Modify, a new video generation model that allows users to modify existing videos by providing character reference images. You can inject yourself everywhere now
-

ChatGPT Images vs NanoBanana: Best Image AI of 2026
By
–
#ChatGPT Images vs #NanoBanana : which is the best Image AI in 2026? I've compared everything in my latest video https://
youtu.be/eEQDlcDz8oQ 10 essential criteria tested and analyzed under the microscope to find out the big winner! -
Photogrammetry from Video Frames using Main Camera
By
–
You can then perform photogrammetry using the captured video frames from the main camera.
-

AI Exceeds Humans in Technical Tasks, Humans Lead in Multimodal Reasoning
By
–
AI now exceeds human performance in most technical tasks — but humans still lead in multimodal reasoning The next frontier isn’t about replacing humans, but amplifying how we think across data, images & language. Source: Stanford AI Index 2025
#AI #Innovation #FutureOfWork -
Gemini 3 Flash Enables Research Paper Analysis and Comparison
By
–
Introducing Gemini 3 Flash for understanding research papers 🚀
— alphaXiv (@askalphaxiv) 17 décembre 2025
Highlight any section of a paper to ask questions and “@” other papers for quick context, comparisons, and benchmark references pic.twitter.com/u3pZc78mn5Introducing Gemini 3 Flash for understanding research papers Highlight any section of a paper to ask questions and “@” other papers for quick context, comparisons, and benchmark references
-
Reddit AMA with SAM 3, SAM 3D, and SAM Audio Researchers
By
–
We’re hosting a Reddit AMA with the researchers behind SAM 3 + SAM 3D + SAM Audio. Join us tomorrow at 2pm PT.