What you can get when you combine deep learning #AI of DXA scans and genomics in >30,000 people! @ScienceMagazine @uk_biobank @ScienceVisuals https://
doi.org/10.1126/scienc
e.adf8009
… https://
science.org/toc/science/cu
rrent
…
MULTIMODAL AI
-

Deep Learning AI Analyzes DXA Scans Genomics 30000 People
By
–
-
Weekly Thread on Latest AI and ML Developments
By
–
In the light of our short context window and lots of cool things that are coming out instantly, I am starting a weekly running thread of things that I find useful spanning a wide variety of topics from LLMs, vision, robotics, multimodals LLMs, etc… I have been jotting down
-
Google Imagen Image Generation Model Now Available on Cloud API
By
–
Nice to see Imagen, Google’s image generation model, finally available on @googlecloud API for external users.
-
DNA-Rendering: Diverse Neural Actor Repository for High-Fidelity Human Rendering
By
–
DNA-Rendering: A Diverse Neural Actor Repository for High-Fidelity Human-centric Rendering
— AK (@_akhaliq) 20 juillet 2023
paper page: https://t.co/Z8tHlgT4EC
Realistic human-centric rendering plays a key role in both computer vision and computer graphics. Rapid progress has been made in the algorithm aspect… pic.twitter.com/WiKQqglDbqDNA-Rendering: A Diverse Neural Actor Repository for High-Fidelity Human-centric Rendering paper page: https://huggingface.co/papers/2307.10173
… Realistic human-centric rendering plays a key role in both computer vision and computer graphics. Rapid progress has been made in the algorithm aspect -

Android in the Wild: Large-Scale Dataset for Device Control
By
–
Android in the Wild: A Large-Scale Dataset for Android Device Control paper page: https://
huggingface.co/papers/2307.10
088
… There is a growing interest in device-control systems that can interpret human natural language instructions and execute them on a digital device by directly controlling -

Unified Agent Framework with Foundation Models for Multimodal AI
By
–
Towards A Unified Agent with Foundation Models paper page: https://
huggingface.co/papers/2307.09
668
… Language Models and Vision Language Models have recently demonstrated unprecedented capabilities in terms of understanding human intentions, reasoning, scene understanding, and planning-like -
AI Generated Animals with Text to Video Using Zeroscope XL
By
–
AI generated Animals with text to video, Zeroscope XL by @IAvadiev pic.twitter.com/EzVbOjWRBV
— AK (@_akhaliq) 19 juillet 2023AI generated Animals with text to video, Zeroscope XL by @IAvadiev
-
Google Cloud PaLM, Imagen, Codey APIs Now Generally Available
By
–
The PaLM language model API, Imagen generative image API, Codey coding APIs, and Chirp speech recognition APIs are now generally available on @googlecloud
! Many have been using them during our public preview, & we're excited they've now hit the GA stage. -
AI Unlocks Animal Communication Mysteries With ML Models
By
–
Ever wondered if we could talk to animals like Dr. Dolittle? AI is unlocking the mysteries of animal communication and our understanding of the animal kingdom with the help of ML and voice-recognition models. Tap to read more: https://
rb.gy/jrekh -
NU-MCC: Multiview Compressive Coding for 3D Reconstruction
By
–
NU-MCC: Multiview Compressive Coding with Neighborhood Decoder and Repulsive UDF
— AK (@_akhaliq) 19 juillet 2023
paper page: https://t.co/fy7pygMD5w
Remarkable progress has been made in 3D reconstruction from single-view RGB-D inputs. MCC is the current state-of-the-art method in this field, which achieves… pic.twitter.com/XFYUp3ddxPNU-MCC: Multiview Compressive Coding with Neighborhood Decoder and Repulsive UDF paper page: https://
huggingface.co/papers/2307.09
112
… Remarkable progress has been made in 3D reconstruction from single-view RGB-D inputs. MCC is the current state-of-the-art method in this field, which achieves