We believe SeamlessM4T represents a significant breakthrough and as part of our open approach, today we're publicly releasing this work under a CC BY-NC 4.0 license so that others can continue to build on this important field of study. Get the code https://
bit.ly/3E4mtVZ
GENERATIVE AI
-
Meta Releases SeamlessM4T Translation Model Openly
By
–
-

SeamlessM4T: All-in-One Multilingual Multimodal Translation Model
By
–
Introducing SeamlessM4T, the first all-in-one, multilingual multimodal translation model. This single model can perform tasks across speech-to-text, speech-to-speech, text-to-text translation & speech recognition for up to 100 languages depending on the task. Details
-
Cursor AI Coding Tool Praised by Developers in Latent Space Podcast
By
–
“Cursor is the best product I've used in a while” – @MacCaw “It's so elegant and easy.” – @AndrewMcCalip “Coding with AI is getting insane.” – @MckayWrigley The Latent Space pod is proud to present: the first podcast with @amanrsanger of @anysphere
! -

Fine-tuning and ControlNet Advances for SDXL Text-to-Image AI
By
–
Seeing the fast follow from the community racing to get fine-tuning and ControlNet working for SDXL has been exciting. These techniques bring text-to-image AI to life. Here are some examples of my dog Queso of SD2.1 Dreambooth (left) vs. SDXL LoRA (right). The quality is
-
Stable Diffusion XL 1.0 Released: Community Discovers Amazing Results
By
–
In July 2023, Stable Diffusion XL 1.0 was released. The 1.0 release of SDXL marks a huge leap forward over previous generations. It has only been a few weeks, and the community is still discovering how to make it amazing. So far the results have been stunning.
-

Stable Diffusion 2: The Awkward Second Album
By
–
In October 2022, Stable Diffusion 2 was released. It was sort of like the awkward 2nd album. Good, just… different. The migration to OpenCLIP as the text encoder significantly changed image composition and limited the use of artists' names that could be used in prompts.
-

Stable Diffusion 1.4 Celebrates One Year Anniversary
By
–
On August 22, 2022 (1yr ago today!) Stable Diffusion 1.4 made its debut. Happy Birthday Stable Diffusion! At last, it seemed like it was finally possible to generate consistently good images from a text prompt.
-

VQGAN and CLIP Notebooks: A Significant Step Forward in Image Generation
By
–
In April 2022, @RiversHaveWings shared a series of Colab notebooks that combined VQGAN and CLIP. https://
x.com/RiversHaveWing
s/status/1516582795438567424
… This was a significant step forward. Images were starting to resemble their prompts, and textures like brush strokes and pencil marks were starting to emerge. -

Pixray reaches 1.3 million runs on Replicate platform
By
–
In early 2022 @dribnet
's Pixray was the first text-to-image model on Replicate to reach thousands of runs. Today it's been run a total of 1.3mil times. @dribnet was actually the first Replicate user to request that we build an API instead of just a web ui. The rest is history -

DALL-E Mini Launch: July 2021 Breakthrough by Boris Dayma
By
–
Fast-forward a few months to July 2021: @borisdayma published DALL·E Mini https://
x.com/borisdayma/sta
tus/1421117516605267968
… This is an excellent breakdown of how it was all put together: https://
wandb.ai/dalle-mini/dal
le-mini/reports/DALL-E-Mini-Explained–Vmlldzo4NjIxODA
…