4. Hugging Face Diffusers We are noticing the recent trend with applications using Diffusion Models either it can be Stable Diffusion or Dalle E Diffusers library provide you with pre trained diffusion models across vision and audio. Check this: https://
github.com/huggingface/di
ffusers
…
CREATIVE AI
-
Hugging Face Diffusers: Pre-trained Diffusion Models for Vision and Audio
By
–
-
Projection Mapping Transforms Kids Dining Experience Into Interactive Art
By
–
How to get kids to be patient and enjoy a dinner at the restaurant 🍲
— Pascal Bornet (@pascal_bornet) 10 février 2023
Would you try it?
Credit: Le Petit Chef, via Lalocal4life#projectionmapping #tech #art #innovation pic.twitter.com/UeHKHMz7WMHow to get kids to be patient and enjoy a dinner at the restaurant Would you try it? Credit: Le Petit Chef, via Lalocal4life
#projectionmapping #tech #art #innovation -
KREA Canvas: AI-Powered Futuristic Coat Design Tool
By
–
"an image is worth a thousand prompts"
— KREA AI (@krea_ai) 10 février 2023
designing futuristic coats with image references in the KREA Canvas ⚡️
we keep sending invites as we can handle more users, sign up here https://t.co/M34dD3yXaw pic.twitter.com/inYGIqjuIB"an image is worth a thousand prompts" designing futuristic coats with image references in the KREA Canvas we keep sending invites as we can handle more users, sign up here https://
forms.gle/L8G6V1pzLx76Pu
R3A
… -
Noise2Music: Transforming Random Noise into Music with AI
By
–
—
Link to project: https://
google-research.github.io/noise2music/ Link to paper: https://
google-research.github.io/noise2music/no
ise2music.pdf
… Authors: Qingqing Huang, Daniel S. Park, Tao Wang, Timo I. Denk, Andy Ly, Nanxin Chen, Zhengdong Zhang, Zhishuai Zhang, Jiahui Yu, Christian Frank, Jesse Engel, Quoc V. Le, William Chan, Wei Han
— -
Zero-shot classification matches text to music without training
By
–
The model knows how to connect text and music to match each of these sentences with a specific music clip, without having to specifically train the model for each clip.
— AI Breakfast (@AiBreakfast) 9 février 2023
This process is called zero-shot classification. pic.twitter.com/KWerrsnMWjThe model knows how to connect text and music to match each of these sentences with a specific music clip, without having to specifically train the model for each clip. This process is called zero-shot classification.
-
Intermediate representation in music generation from text prompts
By
–
"Intermediate representation" refers to a representation of the music generated from the text prompt. The generator model first turns the text prompt into this intermediate representation, which captures elements of the music such as: -Genre
-Tempo
-Instruments
-Mood
-Era -

Noice2Music generates 30-second music from text inputs
By
–
"Noice2Music" is a project from Google Research that generates a short 30 second piece of music based on text inputs.
-
Training generator and cascader models on 300k+ hours of audio
By
–
The 300k+ hours of audio clips were used to train a "generator model" that turns the text into an intermediate representation, and a "cascader model" that uses this intermediate representation to produce high-quality audio.
-
Creation of a training set for Noise2Music with LaMDA
By
–
The researchers created a training set for Noise2Music by using two models to label a collection of 6.8M music source files. They used a large language model (LaMDA in this case) to come up with sentences that describe music in a general way.
-
Future of music creation via text prompts using diffusion models
By
–
In the not-too-distant future, anyone will be able to create any type of music they want via text prompts.
— AI Breakfast (@AiBreakfast) 9 février 2023
Here's a look into recent research on diffusion models for generating high quality music audio from text prompts:
(more examples below) pic.twitter.com/ywgJRVoxLxIn the not-too-distant future, anyone will be able to create any type of music they want via text prompts. Here's a look into recent research on diffusion models for generating high quality music audio from text prompts: (more examples below)