AI Dynamics

Global AI News Aggregator

About

vLLM-Omni: Multi-Modal Framework for Text, Image, Video, Audio

vLLM just got a major upgrade! Originally built for autoregressive text-based LLM serving, vLLM-Omni now lets you serve text, image, video, and audio models – all from a single framework. You can also serve diffusion models for fast parallel generation. 100% open-source.

→ View original post on X — @akshay_pachaar