Thanks for joining us to kick off our first ever Llama track at Connect — lots to share this week and we're just getting started!
GENERATIVE AI
-

JT-CV-9B: High-Quality Text-to-Video Models
By
–
JT-CV-9B
— AK (@_akhaliq) 25 septembre 2024
High-quality Text-to-video Models pic.twitter.com/cmVLmSMSopJT-CV-9B High-quality Text-to-video Models
-

MaskBit: Embedding-Free Image Generation via Bit Tokens
By
–
MaskBit
— AK (@_akhaliq) 25 septembre 2024
Embedding-free Image Generation via Bit Tokens
1. We study the key ingredients of recent closed-source VQGAN tokenizers and develop a publicly available, reproducible, and high-performing VQGAN model, called VQGAN+, achieving a significant improvement of 6.28 rFID over… pic.twitter.com/NoTpp86WDZMaskBit Embedding-free Image Generation via Bit Tokens 1. We study the key ingredients of recent closed-source VQGAN tokenizers and develop a publicly available, reproducible, and high-performing VQGAN model, called VQGAN+, achieving a significant improvement of 6.28 rFID over
-

Alibaba MIMO: Controllable Character Video Synthesis with Spatial Decomposition
By
–
Alibaba presents MIMO
— AK (@_akhaliq) 25 septembre 2024
Controllable Character Video Synthesis with Spatial Decomposed Modeling
Character video synthesis aims to produce realistic videos of animatable characters within lifelike scenes. As a fundamental problem in the computer vision and graphics community, 3D… pic.twitter.com/sAozQvggNzAlibaba presents MIMO Controllable Character Synthesis with Spatial Decomposed Modeling Character video synthesis aims to produce realistic videos of animatable characters within lifelike scenes. As a fundamental problem in the computer vision and graphics community, 3D
-
LLaMA-Omni: Open-Source GPT-4o Alternative from China
By
–
Nice!Here is an opensource GPT-4o from China— LLaMA-Omni.https://t.co/sA1G0uCd0fhttps://t.co/L5jxJRZb5Jhttps://t.co/YiCyg9NiQo
— 机器之心 JIQIZHIXIN (@jiqizhixin) 25 septembre 2024
The work proposes LLaMA-Omni, a novel model architecture designed for low-latency and high-quality speech interaction with LLMs. https://t.co/MOtcT90yThNice!Here is an opensource GPT-4o from China— LLaMA-Omni. https://
arxiv.org/pdf/2409.06666 https://
github.com/ictnlp/LLaMA-O
mni
… https://
huggingface.co/ICTNLP/Llama-3
.1-8B-Omni
… The work proposes LLaMA-Omni, a novel model architecture designed for low-latency and high-quality speech interaction with LLMs. -

Making Text Embedders Few-Shot Learners with In-Context Learning
By
–
Making Text Embedders Few-Shot Learners discuss: https://
huggingface.co/papers/2409.15
700
… Large language models (LLMs) with decoder-only architectures demonstrate remarkable in-context learning (ICL) capabilities. This feature enables them to effectively handle both familiar and novel tasks by -

Global Leaders Discuss AI and Pressing World Issues
By
–
It was an honor to share the stage with @BillGates , Francis Kere, and President @BillClinton to discuss many of the most pressing issues and opportunities of our times, including #AI, at @ClintonGlobal ‘s 2024 Annual Meeting.
