Who out here is training a vector quantised TTS model? Looking for collaborations to bring easy-to-scale training recipes for TTS.
DM?
@reach_vb
-
Vector Quantised TTS Model Training Collaboration Sought
By
–
-
100K Open Source Audio Image Models for Developers
By
–
We have a collection of 100K+ open source models used and loved by millions of developers for Audio and Image. Happy to help bring them closer to devs and unlock plethora of use-cases. Let’s talk!
-

LEDITS: AI Image Editing Tool Now Available on Hugging Face
By
–
I don't see any difference between the two images 🙂 LEDITS by @linoy_tsaban – Try it out now: https://
huggingface.co/spaces/editing
-images/ledits
… -
User expresses fear and hesitation about AI technology adoption
By
–
I am quite scared of using it tbh.
-
Hugging Face Integration Opportunity and Collaboration Proposal
By
–
Wohoo! Really exciting! Keen to see a HF integration.. do let us know how can we help/ collaborate?
-
FastSpeech2 Vocoder Inference Debugging: Tensor Dimension Error
By
–
Running inference on an assortment of FastSpeech2 variants with a new vocoder.. after 3 hours of sifting through the stack trace, realised that I forgot to `unsqueeze` my input vector! *cries in torch*
-

Coqui AI Open Source Text-to-Speech Solution Gains Recognition
By
–
Brilliant! Open source Text-to-Speech ftw! Kudos @coqui_ai
-
Sharing AI Models on Hub Platform
By
–
Fantastic!! Would be cool to get the models on the hub too! Happy to help with it, if you want!
-
Improving Whisper Transcription Quality with Contrastive Search Strategy
By
–
In my experience those results were quite suboptimal and didn’t quite result in usable transcriptions. With better decoding strategy those issues can be alleviated a bit. So the hack is more so of using contrastive search with Whisper to enable such use cases. Will run more
-
Fine-Tuning Whisper Model for Improved Performance
By
–
Definitely yes! You can fine-tune Whisper to boost the model performance: https://
huggingface.co/blog/fine-tune
-whisper
…