Stay tuned for more!! We’re on the mission to enable the next gen audio models!
GENERATIVE AI
-
Encodec: The Key Technology Behind MusicGen Audio Processing
By
–
Not quite, Encodec is part of the pipeline for MusicGen. Encodec helps with converting the audio into discrete codebook representation and back! Not as glamorous but it is the key piece behind making MusicGen as effective as it is
-

HierVL: Hierarchical Video-Language Embedding for Temporal Associations
By
–
HierVL is a novel hierarchical video-language embedding that simultaneously accounts for both long-term and short-term associations. Paper https://
bit.ly/3qSmEk6 7/7 -
Learning Video Representations from Large Language Models
By
–
📺 Learning Video Representations from Large Language Models
— AI at Meta (@AIatMeta) 20 juin 2023
This work repurposed pre-trained LLMs to be conditioned on visual input, and finetune them to create automatic video narrators.
Paper ➡️ https://t.co/OsZt3AnAEM
3/7 pic.twitter.com/DLunz5bcgPLearning Representations from Large Language Models This work repurposed pre-trained LLMs to be conditioned on visual input, and finetune them to create automatic video narrators. Paper https://
bit.ly/3NDCWpR 3/7 -
Movie-Themed Hotel Rooms Reddit Discussion Thread
By
–
reddit thread: https://
reddit.com/r/midjourney/c
omments/14dk95u/movie_themed_hotel_rooms/
… -

Movie Themed Hotel Rooms: Jurassic Park Midjourney AI Art
By
–
Movie Themed Hotel Rooms, midjourney AI 1. Jurassic Park
-

QR Code AI Art Generator with Demo and Discussion
By
–
QR Code AI Art Generator discussion: https://
huggingface.co/spaces/hugging
face-projects/QR-code-AI-art-generator/discussions/4
…
demo: https://
huggingface.co/spaces/hugging
face-projects/QR-code-AI-art-generator
… prompt: A grand city in the year 2100, atmospheric, hyper realistic, 8k, epic composition, cinematic, octane render, artstation landscape vista photography by Carr Clifton & Galen Rowell, -
ANITI publishes a position paper on generative AI systems
By
–
@ANITI_Toulouse publishes a position paper on generative AI systems https://actuia.com/actualite/aniti-publie-un-position-paper-sur-les-systemes-dia-generative/
… #AI #artificialintelligence -
EnCodec Model Now Available in Transformers Library
By
–
Want to train your own Bark/MusicGen-like TTS/TTA models? 👀
— Vaibhav (VB) Srivastav (@reach_vb) 20 juin 2023
The SoTA Encodec model by @MetaAI has now landed in 🤗Transformers!
It supports compression up to 1.5KHz and produces discrete audio representations. ⚡️
Model: https://t.co/Hq8rDHBfjw
Colab: https://t.co/MaWVEAMCXs pic.twitter.com/tDpPAdlYHUWant to train your own Bark/MusicGen-like TTS/TTA models? The SoTA Encodec model by @MetaAI has now landed in Transformers! It supports compression up to 1.5KHz and produces discrete audio representations. Model: https://
huggingface.co/docs/transform
ers/main/en/model_doc/encodec#overview
…
Colab: https://
github.com/Vaibhavs10/not
ebooks/blob/main/use_encodec_w_transformers.ipynb
… -
Transformers Integration Enables Large-Scale Text-to-Speech Music Models
By
–
Okay, but why is it a big deal!? Transformers integration allows us to use any LM and dataset from the ecosystem seamlessly to train Text-to-Speech and Text-to-Music models at scale! More exciting announcements on this front soon! https://
github.com/Vaibhavs10/not
ebooks/blob/main/use_encodec_w_transformers.ipynb
…