AI Dynamics

Global AI News Aggregator

About

LlaMa-Omni: Real-time Speech LLM Model Rivals GPT-4o

Introducing LlaMa-Omni, a new model rivaling GPT-4o for real-time speech interaction with LLMs. This simultaneously generates text and speech directly from speech instructions, with a response latency as low as 226ms. Excited to have the author @Poeroz1204 discussing the work!

→ View original post on X — @askalphaxiv