DeepMind unveiled versatile robot resources, a potential ImageNet moment in robotics. https://
analyticsindiamag.com/5-must-know-ge
neral-purpose-robots/
… #robot #artificialintelligence #innovation #technology @jblefevre60 @3itcom @kalydeoo @ipfconline1 @CurieuxExplorer @Ym78200 @Shi4Tech @smaksked @LaurentAlaus @enilev
MULTIMODAL AI
-
DeepMind’s Versatile Robots: An ImageNet Moment for Robotics
By
–
-

Multimodality and Large Multimodal Models Explained
By
–
New blog post: Multimodality and Large Multimodal Models (LMMs) Being able to work with data of different modalities — e.g. text, images, videos, audio, etc. — is essential for AI to operate in the real world. This post covers multimodal systems in general, including Large
-

SANPO: Egocentric Video Dataset for Outdoor Scene Understanding
By
–
Introducing SANPO, a multi-attribute video dataset for outdoor human egocentric scene understanding composed of both real-world and synthetic data, including depth maps and video panoptic masks with a wide variety of semantic class labels. Read more → https://t.co/bY0ePje0Ss pic.twitter.com/z4ZutNy0xn
— Google AI (@GoogleAI) 10 octobre 2023Introducing SANPO, a multi-attribute video dataset for outdoor human egocentric scene understanding composed of both real-world and synthetic data, including depth maps and video panoptic masks with a wide variety of semantic class labels. Read more → https://
goo.gle/3ZISInU -
ChatGPT Vision Capabilities and AI Wearable Product Review
By
–
In this week’s podcast, @MikeKaput and I share our initial experiences with ChatGPT’s vision capabilities, and it’s safe to say that we were impressed and inspired. We also cover the controversial new AI wearable product from @RewindAI (I try to be objective here, but I am not a
-

Free LLM Course: Zero-to-Hero Training and Fine-Tuning Guide
By
–
Super exciting news!! Along with @towards_AI
, @activeloop and @intel disruptor we just released our LLM free course, an attempt for a « zero-to-hero with LLMs » showing everything about LLMs (train, fine-tune, use RAG…) and the course is multi-modal! We have amazing articles, -
Stable Diffusion XL Workshop with AWS and Bedrock SageMaker
By
–
Today at 9am PST don't forget to join @AWS_Partners and Stability AI for a live demo & workshop to demonstrate how Stable Diffusion XL 1.0 allows professional creatives to produce and implement brilliant content in their works with Amazon Bedrock and Amazon SageMaker
-

Audio conversion tool preserves original speaker voice across languages
By
–
🤯 Convert audio into several languages in minutes
— Ben Tossell (@bentossell) 10 octobre 2023
While preserving the voice of the original speaker.
Creators can convert YouTube, Podcasts, etc and reach a much bigger audience.
Combining multilingual speech synthesis, voice cloning, text & audio processing into one tool; pic.twitter.com/lvSTom4qYlConvert audio into several languages in minutes While preserving the voice of the original speaker. Creators can convert YouTube, Podcasts, etc and reach a much bigger audience. Combining multilingual speech synthesis, voice cloning, text & audio processing into one tool;
-

ChatGPT and DALL-E 3 journaling visualization habit
By
–
New habit: Journal in ChatGPT and ask DALL-E 3 to visualize my entries. It’s incredibly powerful. Makes everything feel more vivid and impactful:
-
Runway Research Addresses Stereotypical Biases in Text-to-Image AI
By
–
Runway Research is proud to share its latest paper, Mitigating Stereotypical Biases in Text to Image Generative Systems. This work is an important step towards better representation for all people. Read the paper and access our open sourced prompts:
-
ChatGPT’s New Vision and Audio Capabilities: Implications for AGI
By
–
Explore the #transformative advancements in #ChatGPT's abilities to "see" and "hear," and delve into the potential #implications these #capabilities have with #AI, its role in daily life, and the broader quest for #Artificial #General #Intelligence. https://
forbes.com/sites/bernardm
arr/2023/10/10/what-do-chatgpts-new-capabilities-really-mean-for-us-all/
…