What if your AI could “see” video like a streaming codec—spending tokens only on the most important moments? Introducing LLaVA-OneVision-2 from Glint Lab, AIM for Health Lab, and MVP Lab. Their secret? Codec-stream tokenization: video is treated as a continuous bit-cost
LLaVA-OneVision-2 tokenizes video like codec, focusing on key moments
By
–
