Sabemos que OpenAI sabe contar hasta tres, así que no sería extraño ver este año un DALL-E 3. Y en realidad, con igualar la calidad de Midjourney, ofrecer el control que se está consiguiendo con Stable Diffusion y ponerlo en la API ya ganarían su lugar.
MULTIMODAL AI
-

NVIDIA AI Technology Simulates Eye Contact in Streaming
By
–
Simulate Eye Contact with #AI!
— Ronald van Loon (@Ronald_vanLoon) 2 mars 2023
What's the best thing about this new concept? Leave your comment!
Thank you for sharing your story, @NVIDIA.#NVIDIAPartner #ArtificialIntelligence #ML #Business #Streaming
Discover New Tech💡First! Sign Up: https://t.co/yiJBVc3fPX pic.twitter.com/CGT65nAtzwSimulate Eye Contact with #AI! What's the best thing about this new concept? Leave your comment! Thank you for sharing your story, @NVIDIA
. #NVIDIAPartner #ArtificialIntelligence #ML #Business #Streaming Discover New TechFirst! Sign Up: http://
bit.ly/3KPSy8L -

VALL-E: Revolutionary Text-to-Speech AI Voice Mimicry
By
–
After ChatGPT and DALL-E, meet VALL-E – the text-to-speech #AI that can mimic anyone’s voice
by @lukekhurst @euronews Go to: https://
buff.ly/3WzDaQh #BigData #MachineLearning cc: @ronald_vanloon @yvesmulkers @ravikikan -
Early Disease Detection Through Embedded Vision Technology in Healthcare
By
–
Imagine a world where diseases can be detected early and patient health can be continuously monitored through embedded vision technology in healthcare. Discover the fascinating potential of this game-changing innovation here: https://
bit.ly/3mlx1ug @VisAiLabs @mahaecon -
Team Achieves Incredible Optimization in New AI Embeddings
By
–
The level of optimization you are achieving is incredible. Same as the optimization of the new Embeddings in December. Congratulations to the team!
-

Kosmos-1: Multimodal LLM for vision tasks and nonverbal reasoning
By
–

The next big leap in useable AI: The ability to understand images (see examples) Microsoft's Kosmos-1, a Multimodal Large Language Model (MLLM) conducts various vision tasks – and suggests MLLMs may be capable of nonverbal reasoning Link to paper: https://
arxiv.org/abs/2302.14045 -

Comprehensive Survey on Pretrained Foundation Models from BERT to ChatGPT
By
–
A Comprehensive Survey on Pretrained Foundation Models: A History from BERT to ChatGPT A nice review of recent advances, current and future researches in pretrained foundation models in NLP, computer vision, graph learning, and other modalities. https://
arxiv.org/abs/2302.09419 -

Stable Diffusion and Midjourney: Image Generation Challenge
By
–
Hay un tema fascinante que me quiero reservar para un futuro vídeo, que querría sacar después de explicar los modelos de difusión. Pero mientras, os comparto el desafío… EXPERTOS EN STABLE DIFFUSION Y/O MIDJOURNEY ¿Podríais hacerme una imagen tal que así?
-

Midjourney AI Revolution: History, Possibilities, and Access Guide
By
–
Midjourney AI has revolutionized the world of art. Check out our latest blog post to know more about its history, possibilities, and how to access this fantastic tool. https://
learnopencv.com/rise-of-midjou
rney-ai-art/
… Check out our Kickstarter: https://
bit.ly/3jQvFqM -

AI-Powered Visual Interfaces: Microsoft’s ChatGPT Integration
By
–
De todos los ejemplos me ha llamado la atención este de aquí. Sobre todo siendo este trabajo de Microsoft. ¿Es evidente, no? Es la predicción que hice respecto a cómo en un futuro nuestras interfaces estarán controladas por IAs como ChatGPT. Ahí la componente visual es CLAVE!