What if your text-to-image model could master every task without trade-offs? Researchers from USTC, UCLA, CUHK, and Xiaohongshu present Flow-OPD. They distill specialized teachers into one student via on-policy learning, plus a regularizer to protect image quality. Result:
MULTIMODAL AI
-
NVIDIA Cosmos 3 launch: download models on HuggingFace and GitHub
By
–
Get started with Cosmos 3: Download Models on @huggingface → https://
huggingface.co/collections/nv
idia/cosmos3
… Customize Models on GitHub → https://
github.com/nvidia/Cosmos -
Cosmos 3: NVIDIA’s open world foundation model for physical AI
By
–
Physical AI needs to understand the world before it can act in it.
— NVIDIA (@nvidia) 4 juin 2026
Introducing Cosmos 3, the open world foundation model and the first omni-model for physical AI.
It understands and generates across text, image, video, sound and action – thanks to a new breakthrough… https://t.co/K5RyHTY7gWPhysical AI needs to understand the world before it can act in it. Introducing Cosmos 3, the open world foundation model and the first omni-model for physical AI. It understands and generates across text, image, video, sound and action – thanks to a new breakthrough
-

ChatLLM Integrates Grok 1.5 Imagine, Auto-Routes to Best AI Model
By
–
ChatLLM now has Grok 1.5 Imagine Use The Best AI For Your Use Case. We smartly route to the best model for you AUTOMATICALLY – SeeDance 2.0, Grok 1.5 imagine
Front-end – Opus 4.8
Back-end – GPT 5.5 xHigh
Chat- Flash 3.5 Cheap – DeepSeek Flash Image – GPT -

New Benchmark for Visual State Tracking in Video Understanding
By
–
"Benchmarking Visual State Tracking in Multimodal Understanding" A new benchmark for tracking visual states. Even though video MLLMs can describe clips really well, they still cannot reliably track what changes over time. This benchmark contains 834 videos and 1,500
-
Gary Marcus offers ‘truly impressed’ award for AI completing Baldur’s Gate 3
By
–
Offering a “damn, I’m truly impressed” award, for the first person or team to build a domain-general AI system that can play @baldursgate3, start to finish.
— Gary Marcus (@GaryMarcus) 4 juin 2026
Will make the prize especially sweet if you manage this decade. pic.twitter.com/M0f5BCh89pOffering a “damn, I’m truly impressed” award, for the first person or team to build a domain-general AI system that can play @baldursgate3
, start to finish. Will make the prize especially sweet if you manage this decade. -
Flows Agent lets you iterate and modify the pipeline dynamically
By
–
Flows Agent lets you iterate through conversation. Tell it to try a warmer voice, swap the background, or generate a version in Spanish. The agent modifies the pipeline and re-runs without rebuilding from scratch.
-
ElevenCreative Flows links 50+ models with voice, music, SFX
By
–
ElevenCreative Flows connects 50+ image and video models with voice, music, and SFX on one canvas. Creators and marketers use it to chain modalities together into complete pipelines and test creative variants across products, languages, and formats.
-
Clippy is back, powered by new Microsoft MAI models for code judgment
By
–
CLIPPY 👏 IS 👏 BACK 👏
— Charly Wargnier (@DataChaz) 4 juin 2026
but this time he’s powered by frontier AI models ready to judge your code 🙈
Microsoft’s new MAI models just dropped on @aimlapi
They recreated Windows XP using MAI-Thinking-1 + @crewAIInc, and brought our fave assistant to life using MAI-Image 2.5 👀↓ https://t.co/AkCibYJJFyCLIPPY IS BACK but this time he’s powered by frontier AI models ready to judge your code Microsoft’s new MAI models just dropped on @aimlapi They recreated Windows XP using MAI-Thinking-1 + @crewAIInc
, and brought our fave assistant to life using MAI-Image 2.5 ↓ -
Krea 2 Turbo: generate high-quality images in 2 seconds
By
–
introducing Krea 2 Turbo.
— Krea (@krea_ai) 4 juin 2026
generate high-quality images in just 2s; compatible with style references, moodboards, and LoRAs.
try it for free at krea . ai pic.twitter.com/cG5wymDdmhintroducing Krea 2 Turbo. generate high-quality images in just 2s; compatible with style references, moodboards, and LoRAs. try it for free at krea . ai
