LibriTTS-R: A Restored Multi-Speaker Text-to-Speech Corpus paper introduces a new speech dataset called “LibriTTS-R'' designed for text-to-speech (TTS) use. It is derived by applying speech restoration to the LibriTTS corpus, which consists of 585 hours of speech data at 24 kHz
GENERATIVE AI
-

Photoshop AI Feature Enables Single-Click Image Manipulation
By
–
Thanks to Photoshop's new AI feature you can go from 'we' to 'me' in just one click." pic.twitter.com/ROoWayoC7P
— Lorenzo Green 〰️ (@mrgreen) 31 mai 2023Thanks to Photoshop's new AI feature you can go from 'we' to 'me' in just one click."
-

AlteredAvatar: Fast Style Adaptation for Dynamic 3D Avatars
By
–
AlteredAvatar: Stylizing Dynamic 3D Avatars with Fast Style Adaptation
— AK (@_akhaliq) 31 mai 2023
presents a method that can quickly adapt dynamic 3D avatars to arbitrary text descriptions of novel styles. Among existing approaches for avatar stylization, direct optimization methods can produce excellent… pic.twitter.com/k9uhlZhWz0AlteredAvatar: Stylizing Dynamic 3D Avatars with Fast Style Adaptation presents a method that can quickly adapt dynamic 3D avatars to arbitrary text descriptions of novel styles. Among existing approaches for avatar stylization, direct optimization methods can produce excellent
-
Midjourney launches billing support email service
By
–
we have a billing support email on our website now billing@midjourney.com pls give them a ring and they should be able to help with any issues
-

Grammar Prompting for Domain-Specific Language Generation with LLMs
By
–
Grammar Prompting for Domain-Specific Language Generation with Large Language Models Large language models (LLMs) can learn to perform a wide range of natural language tasks from just a handful of in-context examples. However, for generating strings from highly structured
-

PaLI-X: Scaling Multilingual Vision and Language Models
By
–
PaLI-X: On Scaling up a Multilingual Vision and Language Model present the training recipe and results of scaling up PaLI-X, a multilingual vision and language model, both in terms of size of the components and the breadth of its training task mixture. Our model achieves new
-

KAFA: Knowledge-Augmented Vision-Language Models for Image Ad Understanding
By
–
KAFA: Rethinking Image Ad Understanding with Knowledge-Augmented Feature Adaptation of Vision-Language Models Image ad understanding is a crucial task with wide real-world applications. Although highly challenging with the involvement of diverse atypical scenes, real-world
-

HiFA: Advanced Diffusion Guidance for High-Fidelity Text-to-3D Synthesis
By
–
HiFA: High-fidelity Text-to-3D with Advanced Diffusion Guidance
— AK (@_akhaliq) 31 mai 2023
Automatic text-to-3D synthesis has achieved remarkable advancements through the optimization of 3D models. Existing methods commonly rely on pre-trained text-to-image generative models, such as diffusion models,… pic.twitter.com/Jr3oJkNzFGHiFA: High-fidelity Text-to-3D with Advanced Diffusion Guidance Automatic text-to-3D synthesis has achieved remarkable advancements through the optimization of 3D models. Existing methods commonly rely on pre-trained text-to-image generative models, such as diffusion models,
-

StyleAvatar3D: High-Fidelity 3D Avatar Generation Using Diffusion Models
By
–
StyleAvatar3D: Leveraging Image-Text Diffusion Models for High-Fidelity 3D Avatar Generation
— AK (@_akhaliq) 31 mai 2023
present a novel method for generating high-quality, stylized 3D avatars that utilizes pre-trained image-text diffusion models for data generation and a Generative Adversarial Network… pic.twitter.com/mYxICiPCgHStyleAvatar3D: Leveraging Image-Text Diffusion Models for High-Fidelity 3D Avatar Generation present a novel method for generating high-quality, stylized 3D avatars that utilizes pre-trained image-text diffusion models for data generation and a Generative Adversarial Network
-
AI E-commerce Demo with Team UI Development Showcase
By
–
T-1 minute to demos. @HowardBGil shows @masadfrost AI app for e-commerce. @ItMeCassie and squad build out a UI for evaluating demos. @pritopian and @giansegato conquered git. (Smiling!)
