When #AI capabilities expand from vision and hearing accuracy, now to smell, as in predicting smell of 500,000 molecules @ScienceMagazine https://
science.org/doi/full/10.11
26/science.ade4401
… https://
science.org/content/articl
e/ai-rivals-the-human-nose-when-it-comes-to-naming-smells
… @epennisi
MULTIMODAL AI
-

AI Now Predicts Smells: 500,000 Molecules Recognition Breakthrough
By
–
-
What Is a Multimodal Model? GPT-4 Explained
By
–
What is a modality or a multimodal model? A multimodal model is a tongue twister. GPT4 is OpenAI's first multimodal model, which just means that it can take multiple modalities, or multiple inputs like text, and image in this case, and in the future, other modalities (eg audio).
-
Meta Releases FACET: Fairness Benchmark for Vision Models
By
–
We’re also releasing FACET, a comprehensive benchmark dataset to evaluate fairness of models across different vision tasks involving people. It consists of 32K images, labeled by expert annotators for demographic & physical attributes.
— AI at Meta (@AIatMeta) 31 août 2023
Explore FACET ➡️ https://t.co/XWXjlY9eV7 pic.twitter.com/a3OtCiNsaSWe’re also releasing FACET, a comprehensive benchmark dataset to evaluate fairness of models across different vision tasks involving people. It consists of 32K images, labeled by expert annotators for demographic & physical attributes. Explore FACET https://
bit.ly/3syYHPy -
DINOv2 License Expansion and FACET Benchmark for Vision Fairness
By
–
Today we’re announcing two new updates in our computer vision work — a new, expanded license for our DINOv2 model and the release of FACET, a comprehensive new benchmark dataset to help evaluate and improve fairness in vision models.
— AI at Meta (@AIatMeta) 31 août 2023
More details ➡️ https://t.co/fDHYNpGrta
🧵 pic.twitter.com/dOXDWOLKSYToday we’re announcing two new updates in our computer vision work — a new, expanded license for our DINOv2 model and the release of FACET, a comprehensive new benchmark dataset to help evaluate and improve fairness in vision models. More details https://
bit.ly/3L35E1U -
Ideogram AI: Generate Images from Text Prompts Easily
By
–
It is @ideogram_ai
. Just enter text prompt and neat things pop out. -
Image Models Limitations: Computational Constraints and Reduced Capability
By
–
Images are hundreds of thousands of pixels, so nobody can afford an architecture that runs a large model over every pixel. The resulting small models are in fact much stupider than GPT-4 and have trouble following even slightly complicated instructions.
-

Gen-2 Motion Slider Feature Released for Video Generation
By
–
We’ve released a new feature in Gen-2: Motion Slider. Select a value from 1 to 10 to control the amount of movement in your output.
— Runway (@runwayml) 30 août 2023
Available now in browser and coming soon to iOS. pic.twitter.com/AZsxkmzAefWe’ve released a new feature in Gen-2: Motion Slider. Select a value from 1 to 10 to control the amount of movement in your output. Available now in browser and coming soon to iOS.
-
Are AI Avatar Businesses Venture Scalable?
By
–
do you think AI avatar businesses are venture scalable businesses?
-
CoTracker: Advanced Pixel Tracking for Video Analysis
By
–
CoTracker can track every pixel in a video, points sampled on a regular grid on any video frame or manually selected points. In our testing it compares favorably against state-of-the-art point tracking methods in both efficiency & accuracy.
— AI at Meta (@AIatMeta) 29 août 2023
Code ➡️ https://t.co/d0igWWYFOu pic.twitter.com/hSO0wbqNu9CoTracker can track every pixel in a video, points sampled on a regular grid on any video frame or manually selected points. In our testing it compares favorably against state-of-the-art point tracking methods in both efficiency & accuracy. Code https://
bit.ly/3KZoLcQ -
CoTracker: Multi-point Video Tracking with Transformer Networks
By
–
New on @huggingface — CoTracker simultaneously tracks the movement of multiple points in videos using a flexible design based on a transformer network — it models correlation of the points in time via specialized attention layers.
— AI at Meta (@AIatMeta) 29 août 2023
🤗 Try CoTracker ➡️ https://t.co/IdaUrH9Xxr pic.twitter.com/WogyQI3F4vNew on @huggingface — CoTracker simultaneously tracks the movement of multiple points in videos using a flexible design based on a transformer network — it models correlation of the points in time via specialized attention layers. Try CoTracker https://
bit.ly/3swQFqt