AI Dynamics

Global AI News Aggregator

About

LLaVA-OneVision-2 tokenizes video like codec, focusing on key moments

What if your AI could “see” video like a streaming codec—spending tokens only on the most important moments? Introducing LLaVA-OneVision-2 from Glint Lab, AIM for Health Lab, and MVP Lab. Their secret? Codec-stream tokenization: video is treated as a continuous bit-cost

→ View original post on X — @jiqizhixin