easy to compare a lot of images from both models on http://
stableboost.ai , e.g. "cute dog cooking tacos, photorrealistic", grid of boosted images from 1.5 (left) and 2.0 (right). 2.0 looking more distorted, cartoony, simpler, ignores text more. may need more prompt engineering
MULTIMODAL AI
-

Comparing Stable Diffusion 1.5 vs 2.0 Image Generation Quality
By
–
-
Stable Diffusion 2.0 Shows Quality Decline Compared to 1.5
By
–
plot twist: stable diffusion 2.0 looks quite a bit worse on the few prompts i've tried so far compared to 1.5 (even not including celebrities/artists). Running theory seems to be this is due to an aggressive data sanitization campaign since the original release (?).
-
Complex Autoencoders Enable Advanced Object Auto-Segmentation in Images
By
–
Fantastic work by @sindy_loewe on complex autoencoders that are able to auto-segment objects in images. https://t.co/RuohaS0WYO
— Max Welling (@wellingmax) 24 novembre 2022Fantastic work by @sindy_loewe on complex autoencoders that are able to auto-segment objects in images.
-
Exploring LAION CLIP Dataset for Keyword Discovery Methods
By
–
is there a way to explore LAION/CLIP so that we can more quickly uncover the key words?
-

Dreambooth Model Training Results with 22 Dog Photos
By
–
I trained dreambooth on just 22 pictures of my pup and I'm amazed by the results!
-

New Fine-Tuned Model Generates Emojis Automatically
By
–
A new fine-tuned model to generate Emojis is out!
-
CLIPSeg: Text-Based Image Inpainting with AI
By
–
Use text to inpaint objects in your images with CLIPSeg.https://t.co/PKtKM57SGN
— KREA AI (@krea_ai) 23 novembre 2022Use text to inpaint objects in your images with CLIPSeg.
-
Free Versatile Diffusion Demo Now Available on Hugging Face
By
–
Free Versatile Diffusion Demo available at @huggingface spaces.https://t.co/SFr2mSRdF4
— KREA AI (@krea_ai) 23 novembre 2022Free Versatile Diffusion Demo available at @huggingface spaces.
-
MagicVideo: Text-to-Video Generation Framework Using Diffusion
By
–
MagicVideo is an efficient text-to-video generation framework based on latent diffusion models that can produce photorealistic video clips from text descriptions.https://t.co/SUEk3csiB7 pic.twitter.com/MqVcOIwNY0
— KREA AI (@krea_ai) 23 novembre 2022MagicVideo is an efficient text-to-video generation framework based on latent diffusion models that can produce photorealistic video clips from text descriptions. https://
magicvideo.github.io -

Versatile Diffusion: Open-Source Multimodal AI Model Released
By
–
Versatile Diffusion can handle text-to-image, image-to-text, image-variation, and text-variation tasks. and the code is open-source! https://
github.com/SHI-Labs/Versa
tile-Diffusion
…