Trending models, datasets, spaces and PAPERS (
https://
huggingface.co/Papers!!) of the week on http://
hf.co. Congrats to @TIIuae (falcon) @MosaicML (mpt) @StabilityAI (SDXl) @MetaAI (MusicGen) @MSFTResearch (Textbooks AAYN) @UMBaltimore (through your eyes) and many others!
MULTIMODAL AI
-

Top AI Models and Datasets Weekly Trends Highlighted
By
–
-
AudioPaLM Multimodal LM Enables Zero-Shot Speech Translation
By
–
10/ AudioPaLM – fuses text-based & speech-based LMs, PaLM-2 and AudioLM, into a multimodal architecture; outperforms existing systems for speech translation tasks and has zero-shot speech-to-text translation capabilities.
-
MotionGPT: Multimodal Control Signals for Human Motion Generation
By
–
8/ MotionGPT – uses multimodal control signals for generating consecutive human motions; it quantizes multimodal control signals intro discrete codes which are converted to LLM instructions that generate motion answers.https://t.co/B2xAYdbuHc
— DAIR.AI (@dair_ai) 25 juin 20238/ MotionGPT – uses multimodal control signals for generating consecutive human motions; it quantizes multimodal control signals intro discrete codes which are converted to LLM instructions that generate motion answers.
-
Futuristic AI Products: When Will They Hit the Market?
By
–
Estos conceptos de productos futuristas no están tan lejos de poder hacerse realidad y llegar al mercado. Estas ideas salen de la mente de @Imkashama y son ciertamente impactantes 😳
— Juan Merodio (@juanmerodio) 25 juin 2023
Te gustaría ver estos productos en el mercado de la mano de estas grandes marcas? pic.twitter.com/3oqhGxMG3EEstos conceptos de productos futuristas no están tan lejos de poder hacerse realidad y llegar al mercado. Estas ideas salen de la mente de @Imkashama y son ciertamente impactantes Te gustaría ver estos productos en el mercado de la mano de estas grandes marcas?
-
Ilumine AI transforms Midjourney images into 3D
By
–
Ilumine AI turning a Midjourney image into 3D pic.twitter.com/iGHH0ukjbb
— AI Breakfast (@AiBreakfast) 25 juin 2023Ilumine AI turning a Midjourney image into 3D
-
Runway Gen-2 Text-to-Video Technology Impresses on Mobile Devices
By
–
I’ve been using @runwayml gen-2 on my phone and it amazes me that t2v can be done period, let alone on a handset.
-

AI Model Draws Thoughts with 80% Accuracy
By
–
Two researchers have created a new #AI model that can draw what you’re thinking with 80% accuracy https://
bit.ly/3FTXOoD -

New Open Source Text to Video AI Model Released
By
–
New open source text to video AI model
— AK (@_akhaliq) 24 juin 2023
576×320 model: https://t.co/fhN2cw2tOn
1024×576: https://t.co/OK7IutR1tF
zeroscope_v2_576w, A watermark-free Modelscope-based video model optimized for producing high-quality 16:9 compositions and a smooth video output. This model was… pic.twitter.com/2w6eYBtUUDNew open source text to video AI model 576×320 model: https://
huggingface.co/cerspense/zero
scope_v2_576w
…
1024×576: https://
huggingface.co/cerspense/zero
scope_v2_XL
… zeroscope_v2_576w, A watermark-free Modelscope-based video model optimized for producing high-quality 16:9 compositions and a smooth video output. This model was -

Zeroscope V2 XL: Watermark-Free High-Quality Video Generation Model
By
–
zeroscope_v2 XL, A watermark-free Modelscope-based video model capable of generating high quality video at 1024 x 576
— AK (@_akhaliq) 24 juin 2023
Model on @huggingface : https://t.co/OK7IutQtE7
This model was trained with offset noise using 9,923 clips and 29,769 tagged frames at 24 frames, 1024×576… pic.twitter.com/K2jJS9N9KBzeroscope_v2 XL, A watermark-free Modelscope-based video model capable of generating high quality video at 1024 x 576 Model on @huggingface : https://
huggingface.co/cerspense/zero
scope_v2_XL
… This model was trained with offset noise using 9,923 clips and 29,769 tagged frames at 24 frames, 1024×576 -
SoundStorm Synthesizes High-Quality Natural Dialogues
By
–
Also, check out this example of how SoundStorm can synthesize high-quality, natural dialogues: https://t.co/vYX5GZ2fJE pic.twitter.com/AkWSLFTSu7
— Google AI (@GoogleAI) 23 juin 2023Also, check out this example of how SoundStorm can synthesize high-quality, natural dialogues: