Stephen Wiltshire is a British artist with autism. He was flown around New York City in a helicopter, and then proceeded to draw the entire city, by memory, on an 19 ft long canvas… Mind blowing! #inspiration
MULTIMODAL AI
-
OpenAI’s Evolution: From API Access to ChatGPT Public Launch
By
–
– Il y a 2 ans OpenAI donnait l'accès aux dev -> création d'app de génération de contenu en surcouche tel que Jasper, rytr, etc. Création de Dall-e, codex…Github copilot – Décembre ouverture de chatGPT au grand public. #AI #autoGPT #AGI
-

Universal Policy Generates Robot Video Trajectories from Images
By
–
Check out a new Universal Policy (UniPi) that, when given an image frame and text description of a task, generates video snippets of what a robot’s trajectory should be and extracts control actions that the robot can execute to achieve that task → https://
goo.gle/41m8F2L -
New Text-to-Video Pipelines and Schedulers Released
By
–
This is just a small subset of things released with the latest release, come check out the wide range of Text-to-Video pipelines, new schedulers and most importantly, our new docs as part of this release!
-
Music Spectrogram Diffusion generates infinite MIDI conditioned music realtime
By
–
Music Spectrogram Diffusion contributed by @krasul
, allows you to generate `infinite` music conditioned via MIDI signal in *realtime* Take it out for a spin noww https://
github.com/Vaibhavs10/not
ebooks/blob/main/text_to_music_with_spectrogram_diffusion_and_diffusers.ipynb
… -
AudioLDM: Text-to-Sound Synthesis with Diffusion Models
By
–
AudioLDM contributed by @sanchitgandhi99, allows you to put in any arbitrary prompt and synthesise high-quality sounds from it. 🔊
— Vaibhav (VB) Srivastav (@reach_vb) 12 avril 2023
Try it out on colab 👉 https://t.co/EhFPgOmrij pic.twitter.com/Jn4nObthSJAudioLDM contributed by @sanchitgandhi99
, allows you to put in any arbitrary prompt and synthesise high-quality sounds from it. Try it out on colab https://
github.com/Vaibhavs10/not
ebooks/blob/main/text_to_sound_with_audioLDM_and_diffusers.ipynb
… -
Diffusers 0.15 brings AudioLDM and Spectrogram Diffusion models
By
–
Diffusers🧨 x Music🎶
— Vaibhav (VB) Srivastav (@reach_vb) 12 avril 2023
Taking diffusers beyond Image ⚡️
With the latest, Diffusers 0.15, we bring two powerful text-to-audio models with all bleeding edge optimisations 💥
1. @LiuHaohe et. al's AudioLDM 🔊
2. @GoogleMagenta's Spectrogram Diffusion 🎹 pic.twitter.com/daU8xW4qcXDiffusers x Music Taking diffusers beyond Image With the latest, Diffusers 0.15, we bring two powerful text-to-audio models with all bleeding edge optimisations 1. @LiuHaohe et. al's AudioLDM 2. @GoogleMagenta
's Spectrogram Diffusion -

Complete toolkit: cloud computing, VR, AR and conversational AI
By
–
Je crois que j'ai tout sous la main, du cloud computing, de la VR, de l'AR, et même de l'agent conversationnel
-
Pix2Video: Text-Guided Video Editing via Image Diffusion
By
–
Adobe & UCL’s Pix2Video: Text-Guided Editing via Image Diffusion Without Preprocessing or Finetuning
-

Deep Learning Predicts Biological Age from Retinal Images
By
–
Images of the retina can provide clues healthcare providers can use to assess a person’s health. In today’s blog, learn how #DeepLearning models can predict biological age from a retinal image and reveal insights that better predict age-related disease → https://
goo.gle/3zNQebs
