Memory works even better on Comet: where we can not only pull details from past threads, but also your open tabs and active projects and workspaces (Google Workspace, Gmail, …) – leading to hyper-personalized results.
MULTIMODAL AI
-

AI Agents Achieve Direct Mind-to-Mind Communication Framework
By
–
AI agents could skip language entirely and communicate mind to mind. This work introduces thought communication, a latent variable framework that identifies shared and private thoughts across agents and recovers the global structure of who shares what. The approach extracts
-

Whisper Thunder challenges VideoGen dominance in text-to-video
By
–
looks like videogen is about to be toppled again. Whisper Thunder ? https://
artificialanalysis.ai/video/leaderbo
ard/text-to-video
… -

20 Must-Try AI Tools for 2025: Complete Stack Guide
By
–
20 must-try AI tools for 2025. This stack covers everything: chat & reasoning image gen video gen voice/music AI search automation AI agents My top 3 right now: ChatGPT, Perplexity, Midjourney. What are yours? #AI #AITools #GenerativeAI #Automation
-

OpenAI Updates Image Generation Model with Improved Quality
By
–
OpenAI quietly updated its image generation model. The improved image quality, the removal of the pee filter, and better physical understanding are impressive, but its multilingual performance and reference consistency still fall short of beating Nano Banana Pro.
-

VLMs Learn Visual Reasoning with COVT Paper
By
–
I just read this paper called "Chain-of-Visual-Thought (COVT)" and it basically teaches VLMs to see and think at the same time not in text, but in continuous visual tokens. Here’s the wild part: Instead of forcing models to reason through words (which destroys all the
-
Predictive AI and Blockchain Transform Future Event Management
By
–
In the future, it could evolve towards predictive AI to personalize crowd flows in real time, AR for immersive interactions, and blockchain for secure and scalable ticketing, elevating events to hyper-connected ecosystems.
-

Vision Language Models Parse Floor Plan Maps Successfully
By
–
Vision Language Models Can Parse Floor Plan Maps! Paper: https://
arxiv.org/abs/2409.12842
Site: https://
sites.google.com/view/vlm-floor
plan/
… -
SAM 3D Advances Rehabilitation Through Human Movement Analysis
By
–
SAM 3D is helping advance the future of rehabilitation.
— AI at Meta (@AIatMeta) 25 novembre 2025
See how researchers at @CarnegieMellon are using SAM 3D to capture and analyze human movement in clinical settings, opening the doors to personalized, data-driven insights in the recovery process.
🔗 Learn more about SAM… pic.twitter.com/UkQyD2qLeVSAM 3D is helping advance the future of rehabilitation. See how researchers at @CarnegieMellon are using SAM 3D to capture and analyze human movement in clinical settings, opening the doors to personalized, data-driven insights in the recovery process. Learn more about SAM
-

Flux.2 Image Model Live on Poe Platform
By
–
Flux.2 is live on Poe! Black Forest Labs’ new state-of-the-art image model delivers sharp detail with strong style control and text rendering. Powered by @fal on Poe, try it for concept art, product shots, and high-fidelity edits. (1/2)
