Chatting with YouTube videos is a Perplexity Comet use case I've been spamming It's an AI tutor for every educational video Whenever I don't fully understand something, I pause the video and do a mini rabbit hole session I've been doing this while reading with ChatGPT Voice
MULTIMODAL AI
-
Apple Launches Multilingual Foundation Models for On-Device AI
By
–
10. Apple Intelligence Foundation Language Models Apple introduces two multilingual, multimodal foundation models: a 3B-parameter on-device model optimized for Apple silicon and a scalable server model using a novel PT-MoE transformer architecture.
-
Best AI Models Capabilities Six Months Ago
By
–
The best models could do 6 months ago. pic.twitter.com/Kg5iAczkaS
— Ethan Mollick (@emollick) 27 juillet 2025The best models could do 6 months ago.
-

WordPress categories related to AI topics
By
–
Admittedly not the best example—just the first that popped up in search. Here’s one from less than a month ago, so presumably using 4o or better:
-
AI Generated Videos: Windows to Projected Future?
By
–
Estos vídeos son sorprendentes. Crees que pueden ser una ventana al futuro proyectada por la inteligencia artificial? pic.twitter.com/E86bfouawp
— Juan Merodio (@juanmerodio) 27 juillet 2025Estos vídeos son sorprendentes. Crees que pueden ser una ventana al futuro proyectada por la inteligencia artificial?
-
AI Enhances Meetings While Preserving Human Empathy and Communication
By
–
This could unlock not only more productive meetings, but also more human ones, where empathy, subtle cues, and non-verbal communication are no longer lost in translation.
-
Hunyuan3D World Model 1.0 Released
By
–
BREAKING 🚨: Hunyuan released its open-source Hunyuan3D World Model 1.0.
— 🚨 AI News | TestingCatalog (@testingcatalog) 27 juillet 2025
It can generate 3D worlds from a single prompt 🤯
This sample is stunning 👀 https://t.co/orwWjUfmtI pic.twitter.com/8nxD05XoruBREAKING : Hunyuan released its open-source Hunyuan3D World Model 1.0. It can generate 3D worlds from a single prompt This sample is stunning
-
Runway Aleph: Advanced In-Context Video Generation Model Launch
By
–
The introduction of Runway Aleph, our state-of-the-art in-context video model setting a new frontier for multi-task visual generation. Chat mode now available on mobile. And our weekly community spotlight featuring incredible Act-Two creations. Get caught up on what happened This… pic.twitter.com/pGcNZJOR8V
— Runway (@runwayml) 27 juillet 2025The introduction of Runway Aleph, our state-of-the-art in-context video model setting a new frontier for multi-task visual generation. Chat mode now available on mobile. And our weekly community spotlight featuring incredible Act-Two creations. Get caught up on what happened This
-

Building a Voice-to-Diagram Note App with Kimi K2 and Groq
By
–
kimi k2 @Kimi_Moonshot + @GroqInc vibe coding one shot create a gradio note taking app that generates diagrams as you speak using Whisper + flux with @FAL + @huggingface inference providers, docs in prompt
-

Zenith vs Kingfall Comparison on SVG Robot
By
–



Zenith (GPT-5?) vs Kingfall (Gemini 3?) SVG Robot is the ultimate benchmark.
