Photo by AI, video by AI, text by AI:
MULTIMODAL AI
-
Thales Launches Multimodal Biometric Module for Facial and Iris Recognition
By
–
The new multimodal biometric module from @thalesgroup combines facial and iris recognition https://actuia.com/actualite/le-nouveau-module-biometrique-multimodal-de-thales-combine-reconnaissance-faciale-et-de-liris/
… #AI #artificialintelligence #security -
Face Recognition Model Quality Issues Pre-Tuning
By
–
Pic example of how bad it is at faces pre-tuning
-
SD Resolution Limits and Future Improvements for Image Quality
By
–
Also a lot is it due to 512×512 limit of SD1.5, with SD2 u can do 786×786 so it gets better, but we're all upscaling so real details we will get in a few months when we can do 1024×1024 and 2048×2048 etc after, it will exponentially get better
-

Avatar AI Achieves Photorealistic Progress in Daily Updates
By
–
Daily http://
avatarai.me photorealistic progress pics 2 -

Avatar AI Achieves Photorealistic Progress in Daily Updates
By
–
Daily http://
avatarai.me photorealistic progress pics -

Robotics Transformer 1: Multi-Task Robot Learning Model
By
–
Introducing the Robotics Transformer 1, a multi-task model that tokenizes robot inputs and outputs actions to enable efficient inference at runtime. Learn how it improves zero-shot generalization to new tasks, environments and objects → https://t.co/hnsKvCJjmP pic.twitter.com/g9QFXzjs9T
— Google AI (@GoogleAI) 13 décembre 2022Introducing the Robotics Transformer 1, a multi-task model that tokenizes robot inputs and outputs actions to enable efficient inference at runtime. Learn how it improves zero-shot generalization to new tasks, environments and objects → https://
goo.gle/3Yxomnt -

Meta AI Unveils data2vec 2.0: 16x Faster Self-Supervised Learning
By
–
Announcing data2vec 2.0, a new general self-supervised algorithm built by Meta AI for speech, vision & text that can train models 16x faster than the most popular existing algorithm for images while achieving the same accuracy. Read more & get the open source code
-

MAGVIT: Masked Generative Video Transformer Explained
By
–
MAGVIT: Masked Generative Transformer Yu et al.: https://
arxiv.org/abs/2212.05199 #Artificialintelligence #DeepLearning #Machinelearning
