Wait wtf!? @GoogleDeepMind Gemini 1.5Pro out scoring @OpenAI O1-preview on FrontierMath :O Even @AnthropicAI 3.5 Sonnet (new) beats it!
@reach_vb
-
GPT-4o Advanced Voice Mode Launch Announcement
By
–
Open GPT-4o Advanced Voice Mode ⚡️ https://t.co/Yae2nFv3gF
— Vaibhav (VB) Srivastav (@reach_vb) 5 novembre 2024Open GPT-4o Advanced Voice Mode
-
Fish Agent v0.1: New Multilingual Speech-to-Speech Model Released
By
–
Wow! New Speech to Speech model – Fish Agent v0.1 3B by @FishAudio 🔥
— Vaibhav (VB) Srivastav (@reach_vb) 5 novembre 2024
> Trained on 700K hours of multilingual audio
> Continue-pretrained version of Qwen-2.5-3B-Instruct for 200B audio & text tokens
> Zero-shot voice cloning
> Text + audio input/ Audio output
> Ultra-fast… pic.twitter.com/UvdwxGUm4wWow! New Speech to Speech model – Fish Agent v0.1 3B by @FishAudio > Trained on 700K hours of multilingual audio
> Continue-pretrained version of Qwen-2.5-3B-Instruct for 200B audio & text tokens
> Zero-shot voice cloning
> Text + audio input/ Audio output
> Ultra-fast -
Using PEFT Adapters in Your Machine Learning Projects
By
–
If you have a PEFT adapter, then you can use:
-
Tencent Releases Hunyuan Large Language Model
By
–
https://
huggingface.co/tencent/Tencen
t-Hunyuan-Large
… -

Tencent Hunyuan Large 389B: New LLM Outperforms Llama DeepSeek
By
–
We're sooo back! – Tencent Hunyuan Large – 389B (Total) X 52B (Active) – beats Llama 3.1 405B, Mistral 8x22B, DeepSeek V2! Multilingual, 128K context, Utilizes GQA + CLA for KV Cache compression + Higher throughput Released Pre-train, Instruct & FP8 checkpoints on the Hugging
-
Invitation to discuss HF collaboration for model iterations
By
–
Thank you! I'd love to chat more and discuss how HF can help with futher iterations of the model, would you mind sending a DM!
-

350M Text-to-Speech Model Generates Impressive Coherent Audio
By
–
It's hilarious to see the model do go off the rails and just make random but coherent audio up – it's still quite impressive for a 350M Text to Speech model 👀 https://t.co/bqsULmSPlK pic.twitter.com/TMJWaHY2KS
— Vaibhav (VB) Srivastav (@reach_vb) 4 novembre 2024It's hilarious to see the model do go off the rails and just make random but coherent audio up – it's still quite impressive for a 350M Text to Speech model
