Introducing LlaMa-Omni, a new model rivaling GPT-4o for real-time speech interaction with LLMs. This simultaneously generates text and speech directly from speech instructions, with a response latency as low as 226ms. Excited to have the author @Poeroz1204 discussing the work!
LlaMa-Omni: Real-time Speech LLM Model Rivals GPT-4o
By
–
