Hyperbolic Image-Text Representations https://
bit.ly/3YlLC8h 5/7
MULTIMODAL AI
-
Hyperbolic Image-Text Representations Research
By
–
-
Multimodal AI for Brand Visual Asset Analysis and Rebranding
By
–
Now I’m curious if anyone has used multimodal capabilities with generative AI to examine visual assets for a potential brand or rebrand.
-

Gen-2 Image to Video Mode: New Creative AI Experiments
By
–
Photo experiments with the new Image to Video mode of Gen-2. pic.twitter.com/ZYebBRQnxB
— Runway (@runwayml) 25 juillet 2023Photo experiments with the new Image to mode of Gen-2.
-

Language Models Guide Advanced Robotic Control Systems
By
–
@SereactAI – more on language model-guided robotic control systems.
-

Google Presents ML Research at ICML Conference Booth
By
–
Attending @icmlconf
? Drop by the Google booth to chat with researchers and learn about our exciting #ML research that spans topics from vision and language models to algorithms and theory. Learn more about our involvement at #ICML2023: https://
goo.gle/44Rxc29 -

Vit-22b and Flan Collection Presented at ICML Conference
By
–
Vit-22b and flan collection are presented at @icmlconf
! -
Multimodal LLM with Vision-Language Transformer and LangChain
By
–
Using the `Vision-and-Language Transformer` model and @LangChainAI to create a Multimodal LLM in @Streamlit! 🔥
— Charly Wargnier (@DataChaz) 23 juillet 2023
– Demo app: https://t.co/miwwtiUxv0
– ViLT model: https://t.co/nz8npLUs9o
– App creator: @nicolas_tch pic.twitter.com/iTWuMwgc8WUsing the `Vision-and-Language Transformer` model and @langchain to create a Multimodal LLM in @Streamlit
! – Demo app: https://
vilt-gpt-ppn83ly4c9.streamlit.app
– ViLT model: https://
huggingface.co/dandelin/vilt-
b32-finetuned-vqa
…
– App creator: @nicolas_tch -

Computer Vision Embeddings for Machine Learning Applications
By
–
Computer Vision Embeddings for Machine Learning https://
bit.ly/3CzwZE7 #AI #MachineLearning #DeepLearning #LLMs #DataScience -
From Links to Answers: AI-Driven Task Completion Evolution
By
–
Accelerate the transition from links –> answers, sifting –> learning, browsing –> getting things done.
-

Meta-Transformer: Unified Learning Across 12 AI Modalities
By
–
8/ Meta-Transformer – performs unified learning across 12 modalities; supports tasks like fundamental perception (text, image, point cloud, audio, video), application (X-Ray, infrared, hyperspectral, & IMU), and data mining (graph, tabular, & time-series)