Using image inputs with GPT-4! Will be available via our API.
MULTIMODAL AI
-

ChatGPT Vision: everything you can do
By
–
That's it! My #ChatGPT has eyes and it's just INCREDIBLE! In this video, I present to you ChatGPT Vision and everything you can do —> https://youtu.be/Ao0EhzoX8kU #GPT4 #ChatGPTV #ChatGPTPlus
-

Runway ML Studios Premieres AI-Generated Film From Tweet
By
–
Last week we asked you to help us generate a film with a Tweet. Today, @runwaymlstudios is excited to premiere, "Don't Pick Up." Thank you to all our collaborators. pic.twitter.com/qUj5JIeWBW
— Runway (@runwayml) 12 octobre 2023Last week we asked you to help us generate a film with a Tweet. Today, @runwaymlstudios is excited to premiere, "Don't Pick Up." Thank you to all our collaborators.
-

Multimodal Evolution of Vector Embeddings in AI
By
–
The Multimodal Evolution of Vector Embeddings https://
bit.ly/3PIx1AJ
#AI #MachineLearning #DeepLearning #LLMs #DataScience -
Legal implications of curated image datasets in AI training
By
–
And how about the "aesthetic" subsets? Are they likely to be treated differently because they're curated? (And they would have required accessing the image itself to do that curation.)
-
Coqui XTTS: Open Source Text-to-Speech with Voice Cloning
By
–
18 – CoquiTTS (
@coqui_ai
) – @_josh_meyer_ and the Coqui team presented XTTS: A fully open source text to speech and voice generation model competitive with the commercial offerings, including instant-voice cloning http://
hf.co/spaces/coqui/X
TTS
… -
HelloHola: AI Speech Translation with Lip Sync
By
–
15 – HelloHola presented by @gary_Offbeat
, it translates speech into a chosen language, maintaining your style and synchronizing lip movements, check them out @ -
DINOv2: State-of-the-art open source self-supervised vision transformer
By
–
9 – DINOv2 – presented by @TimDarcet
, DinoV2 is state of the art open source self supervised vision transformer model https://
dinov2.metademolab.com -

YOLO SAM Combines Meta’s Segment Anything with Medical Imaging
By
–
5 – Sumit Pandey showcased YOLO SAM, a pipeline that combined Segment Anything from @Meta with Medical Image Segmentation for X-Rays, CT Scans and Ultrasounds https://
mlshots.live/YOLO-SAM/softw
are
… -

GPT-4 dominates AI model competition across all metrics
By
–
GPT4 wins the competition on all dimensions (not counting speed and price) but there are some notable callouts in 2nd place. (2/7)