Because 1T parameters might be a bit too much, I'm also releasing an "intermediate checkpoint": BigLlama-3.1-681B-Instruct. It's closer to what Llama 3 120B was to the 70B, so it might be more coherent as well. Model: https://
huggingface.co/mlabonne/BigLl
ama-3.1-681B-Instruct
…
LLMS
-

BigLlama-3.1-681B-Instruct: New Intermediate Checkpoint Released
By
–
-

BigLlama-3.1 reaches 1 trillion parameters milestone
By
–
BigLlama-3.1-1T-Instruct So I've heard that 405B parameters weren't enough… It's my pleasure to present an upscaled Llama 3.1 with 1,000,000,000 parameters. Now available on @huggingface
. Model: https://
huggingface.co/mlabonne/BigLl
ama-3.1-1T-Instruct
… -
Chain of Density Paper Overview and Research Insights
By
–
Check out the chain of density paper from @GriffinAdams92
-
OpenAI GPT-6 training scale and deployment timeline
By
–
Esto es como cuando te tomas un break para el café mientras se entrena tu modelo, pero claro… a la escala de OpenAI. Lanzas el entrenamiento de GPT-6 y te tomas unos meses sabáticos hasta que termine.
-

MiniCPM-V: GPT-4V Level MLLM for Mobile Devices
By
–
MiniCPM-V A GPT-4V Level MLLM on Your Phone paper page: https://
huggingface.co/papers/2408.01
800
… The recent surge of Multimodal Large Language Models (MLLMs) has fundamentally reshaped the landscape of AI research and industry, shedding light on a promising path toward the next AI milestone. -

Lumina-mGPT: Multimodal Generative Pretraining for Text-to-Image
By
–
Lumina-mGPT Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining paper page: https://
huggingface.co/papers/2408.02
657
… We present Lumina-mGPT, a family of multimodal autoregressive models capable of various vision and language tasks, particularly -

Language Model Can Listen While Speaking
By
–
Language Model Can Listen While Speaking https://
huggingface.co/papers/2408.02
622
… Dialogue serves as the most natural manner of human-computer interaction (HCI). Recent advancements in speech language models (SLM) have significantly enhanced speech-based conversational AI. However, these models -

Language Models Can Listen While Speaking Simultaneously
By
–
Language Model Can Listen While Speaking https://
huggingface.co/papers/2408.02
622
… Dialogue serves as the most natural manner of human-computer interaction (HCI). Recent advancements in speech language models (SLM) have significantly enhanced speech-based conversational AI. However, these models -

Language Models Can Listen While Speaking in Dialogue
By
–
Language Model Can Listen While Speaking https://
huggingface.co/papers/2408.02
622
… Dialogue serves as the most natural manner of human-computer interaction (HCI). Recent advancements in speech language models (SLM) have significantly enhanced speech-based conversational AI. However, these models -

Language Models Can Listen While Speaking Simultaneously
By
–
Language Model Can Listen While Speaking https://
huggingface.co/papers/2408.02
622
… Dialogue serves as the most natural manner of human-computer interaction (HCI). Recent advancements in speech language models (SLM) have significantly enhanced speech-based conversational AI. However, these models