I tried making images of popular cartoon characters via DALL-E 3 and it generated some that looked very similar to the actual characters and others with less resemblance. Here’s a few from the other day. Prompts in the ALT text.
MULTIMODAL AI
-

Gemini vs GPT-4V: Vision-Language Models Comparison
By
–
10/ Gemini vs. GPT-4V – a comparison of vision-language models like Gemini & GPT-4V; finds that GPT-4V is precise and succinct in responses, while Gemini excels in providing detailed, expansive answers accompanied by relevant imagery and links.
-
Innovative Multimodal Models Redefine AI Capabilities
By
–
7. Innovative multimodal Models • LLAMA 2, SAM, LLaVA, and more redefine image segmentation, language vision, and advanced Al capabilities.
-
PaLM-E: AI Robotics Breakthrough in Multimodal Interpretation
By
–
5. AI in Robotics with PaLM-E • PaLM-E blends image, text, and robotics, showcasing Al's capability to interpret and interact with the real world.
-
Gen1: Breakthrough in AI-Driven Video Editing Technology
By
–
4. Evolution in Editing • Gen1 based on Stable Diffusion marks a leap in Al-driven video editing – intelligently adapting specific elements.
-
InstructPix2Pix: Text-Driven Image Editing Technology
By
–
2. Text-Driven Image Edits • InstructPix2Pix enables precise image modifications through text instructions, merging text and image seamlessly.
-
MusicLM: AI Generates Music from Text
By
–
3. Innovations in Music Composition • MusicLM generates musical pieces from text, showcasing AI's expanding role in creative domains.
-
Voice Synthesis Revolution: 3-Second Sample Voice Imitation Technology
By
–
1. Voice Synthesis Revolution Valley's technology imitates voices with a 3-second sample Transforming digital media and sound synthesis.
-

Multimodal RAG with LangChain and GPT-4 Vision Guide
By
–
Multimodal RAG using Langchain Expression Language And GPT4-Vision Great guide from Plaban Nayak on multimodal RAG. We suspect this will become more and more popular – so now's a great time to become an expert! Blog: https://
medium.aiplanet.com/multimodal-rag
-using-langchain-expression-language-and-gpt4-vision-8a94c8b02d21
… -

Telescope Magazine Issue 1: Art Film and Human Creativity
By
–
Issue No1 of TELESCOPE MAGAZINE: art, film and human creativity.https://t.co/eB5qXmjurk pic.twitter.com/e11rZTpa2p
— Runway (@runwayml) 30 décembre 2023Issue No1 of TELESCOPE MAGAZINE: art, film and human creativity. http://
TELESCOPEMAGAZINE.com
