MICROSOFT KOSMOS-1 Una semana después de presentaros en el canal la idea de Multimodalidad, vemos un nuevo avance en esta línea por parte de Microsoft. Han entrenado a un modelo que percibe y combina información textual y vision/audio. Imaginad un ChatGPT, que vea y oiga!
MULTIMODAL AI
-
Temporal consistency in diffusion models: evidence and validation
By
–
Send me a link to some paper where they demonstrate temporal consistency of this quality using diffusion models like Stable Diffusion, and I'll believe you.
-
Deep Learning Breakthrough: 3D Avatar Reconstruction from 2D Video
By
–
Otro ejemplo de Deep Learning que me sorprende hoy por su robustez respecto a lo que teníamos hasta hace no mucho. Pensad que lo que se muestra en este vídeo es la reconstrucción del avatar 3D de lo que se observa en el vídeo 2D ¡Flipante! https://t.co/nDHME6vY4x
— Carlos Santana (@DotCSV) 27 février 2023Otro ejemplo de Deep Learning que me sorprende hoy por su robustez respecto a lo que teníamos hasta hace no mucho. Pensad que lo que se muestra en este vídeo es la reconstrucción del avatar 3D de lo que se observa en el vídeo 2D ¡Flipante!
-
Ludwig 0.7 Released with Computer Vision and LLM Improvements
By
–
Ludwig 0.7 is live! The new release is packed with features incl: Expanded support for pretrained #computervision models Fully customizable image augmentation pipelines 50x faster tuning of #largelanguagemodels and more! https://
pbase.ai/Ludwig07 #opensource #ML -

Deep Learning and AI to Transform the Metaverse in 2023
By
–
How #DeepLearning will ignite the metaverse in 2023 and beyond
by Victor Dey @VentureBeat Read more: https://
buff.ly/3PRaRLq #AI #IoT #BigData #MachineLearning #ArtificialIntelligence #ML #MI #AR cc: @ronald_vanloon @yvesmulkers @kuriharan -
AI-Powered Brain Implant Speed Record, Google Image Generator, China Adoption
By
–
Today's edition of The Newsletter: AI-Powered Brain Implant Smashes Speed Record for Turning Thoughts Into Text
Google might bring AI text-to-image generator to Android
China’s AI Adoption
Best AI Tool of the Week
Calling all Engineers -
AR Filter Technology Advances Overcome Previous Tracking Limitations
By
–
Aquí otro ejemplo de a qué me refiero. Antes, pasar la mano por tu cara era suficiente para romper el filtro y ver las pestañas y pintalabios volando por toda tu cara.
— Carlos Santana (@DotCSV) 27 février 2023
Y parece que eso ya es cosa del pasado.https://t.co/w6zLH6160eAquí otro ejemplo de a qué me refiero. Antes, pasar la mano por tu cara era suficiente para romper el filtro y ver las pestañas y pintalabios volando por toda tu cara. Y parece que eso ya es cosa del pasado.
-
AI Facial Filters Break Through Hand Detection Limitations
By
–
En lo técnico os tengo que decir que es impresionante ver lo robusto que es este filtro, donde una mano frente a la cámara o el pelo ya no rompe la ilusión como pasaba anteriormente
— Carlos Santana (@DotCSV) 27 février 2023
Hoy son filtros de maquillaje, mañana será cualquier cara que quieras 😐https://t.co/4JTzv8k24xEn lo técnico os tengo que decir que es impresionante ver lo robusto que es este filtro, donde una mano frente a la cámara o el pelo ya no rompe la ilusión como pasaba anteriormente Hoy son filtros de maquillaje, mañana será cualquier cara que quieras
-

Trained LLM plus image2image with human fine-tuning feasible soon
By
–
This post may or may not be true – but nothing described in it is unrealistic. A trained LLM + image2image. Add some Human-in-the-loop fine-tuning and this project could exist in a few months.
-
Creative Technology Blending Reality and Imagination
By
–
Wielding creative technology to blend reality and imagination
