You just pass a reference speaker audio and it synthesises audio in that persona.
TOOLS
-
Real-time API implementation deployable on L4 under one dollar
By
–
Emphasis on cascaded. We do have real-time API based implementation which you can deploy on a L4 at less than 1$ an hour.
-
Hugging Chat: Exploring Open Source AI Conversational Platform
By
–
That’s great, have you heard about http://
hugging.chat -
Hugging Face Enterprise Hub Workshop with Co-founder CEO
By
–
Live workshop with Hugging Face co-founder and CEO: Step up your AI builder game with Hugging Face’s enterprise hub https://
x.com/i/broadcasts/1
lPKqOyXQnNJb
… -
Perplexity Launches Supply: AI-Focused Product Line
By
–
Introducing Perplexity Supply.
— Perplexity (@perplexity_ai) 30 octobre 2024
Quality goods, thoughtfully designed for curious minds.https://t.co/o32T8Hm6o9 pic.twitter.com/MewR8n4dpBIntroducing Perplexity Supply. Quality goods, thoughtfully designed for curious minds. http://
perplexity.supply -
Gradio WebRTC Conversational AI GitHub Repository
By
–
github: https://
github.com/freddyaboulton
/gradio-webrtc?tab=readme-ov-file#conversational-ai
… -

Build ChatGPT-like Voice Apps with Gradio in Few Lines
By
–
make apps like chatgpt voice in a few lines of code with gradio
— AK (@_akhaliq) 30 octobre 2024
check out this example with omni mini by @freddy_alfonso_
chatgpt advance voice mode asks for 3 places to order pizza and omni mini responds in a gradio UI pic.twitter.com/W2rvQSYujwmake apps like chatgpt voice in a few lines of code with gradio check out this example with omni mini by @freddy_alfonso_ chatgpt advance voice mode asks for 3 places to order pizza and omni mini responds in a gradio UI
-
NotebookLM Technical Deep Dive Explained
By
–
Great technical details on how NotebookLM works. https://t.co/sBv6h2TFFe
— Paul Roetzer (@paulroetzer) 30 octobre 2024Great technical details on how NotebookLM works.
-
Mistral AI and Qualcomm Integrate Mistrals at Edge Devices
By
–
[#Article] Mistral AI and Qualcomm collaborate to integrate Mistrals into edge devices https://actuia.com/actualite/mistral-ai-et-qualcomm-collaborent-pour-integrer-les-ministraux-aux-appareils-en-peripherie/
…
#AI #artificialintelligence -
MaskGCT Open Source Text-to-Speech Model Achieves New SoTA
By
–
Fuck yeah! MaskGCT – New open SoTA Text to Speech model! 🔥
— Vaibhav (VB) Srivastav (@reach_vb) 30 octobre 2024
> Zero-shot voice cloning
> Emotional TTS
> Trained on 100K hours of data
> Long form synthesis
> Variable speed synthesis
> Bilingual – Chinese & English
> Available on Hugging Face
Fully non-autoregressive… pic.twitter.com/CAUX6cTiAGFuck yeah! MaskGCT – New open SoTA Text to Speech model! > Zero-shot voice cloning
> Emotional TTS
> Trained on 100K hours of data
> Long form synthesis
> Variable speed synthesis
> Bilingual – Chinese & English
> Available on Hugging Face Fully non-autoregressive