Jazykovy model je uz take male AGI. Dobre rozumie nasmu svetu, vie riesit problemy atd. Co mu chyba (da sa doplnit) je autonomia, schopnost sa ucit zo skusenosti, dalsie modality (video, real fyzika), schopnost pouzivat externe nastroje, komunikacia s inymi jazykovymi modelmi.
MULTIMODAL AI
-
Moûsai: Cascading Latent Diffusion for High-Quality Text-to-Music
By
–
A new AI research presents a cascading latent diffusion approach called Moûsai that simplifies high-quality text-to-music at 48kHz. To read how this method can generate minutes of high-quality music in real-time, visit: https://
bit.ly/3JPoyt6 @jeevprabnivash @_DigitalIndia -
Outpainting Technology Receives Major Upgrade
By
–
outpainting just got upgraded! ⚡️ pic.twitter.com/mSgttVVygi
— KREA AI (@krea_ai) 8 février 2023outpainting just got upgraded!
-
Dreamix: Diffusion Model for General Video Editing
By
–
Google & HUJI Present Dreamix: The First Diffusion Model for General Editing https://
syncedreview.com/2023/02/07/goo
gle-huji-present-dreamix-the-first-diffusion-model-for-general-video-editing/
… -
Natural voice generation with emotion and age variation
By
–
Here I am talking about natural voice generation. Imagine I say 'say hello in the tone of a 56-year-old man who is sad'. And 'say hello in the tone of a 56-year-old man who is happy'. It's the same word, but not at all the same generation. It's very complex.
-
Future of Generative AI: Text to Voice Generation Timeline
By
–
The future of #GenerativeAI looks cool. Text generation: (2021)
Image generation: (2022) generation: (2023)
Voice generation: (2024) -
Universal Speech Translator Enables Hokkien Speech-to-Speech Translation
By
–
Our Universal Speech Translator (UST) is the first AI-powered translation system providing speech-to-speech translation for Hokkien, one of ~3k primarily spoken languages with no standard writing system & very few human translators.
— AI at Meta (@AIatMeta) 6 février 2023
More on this work ➡️ https://t.co/NwH7zditOD pic.twitter.com/XTAqpFblF1Our Universal Speech Translator (UST) is the first AI-powered translation system providing speech-to-speech translation for Hokkien, one of ~3k primarily spoken languages with no standard writing system & very few human translators. More on this work https://
bit.ly/3Rw3SIl -
Google Announces Bard AI and Search Updates
By
–
Source: https://
blog.google/technology/ai/
bard-google-ai-search-updates/
… -
Elf-Attention Models and the ORC Challenge
By
–
Can models based on the Elf-Attention mechanism tackle the ORC challenge?
-
AI Version of Asmongold Interviewed Live on Stream
By
–
Witnessing history. @AtheneLOL is interviewing in a lifestream a AI version of Asmongold (a very popular streamer) This kind of tech will hit mainstream hard.