Mathematicians, unite! What are your go-to problems for testing the accuracy of LLMs?
@01ai_yi
-

Yi-1.5-34B-Chat Ranks High on Chatbot Arena Leaderboard
By
–
Thank you @lmsyorg! So exciting to see that Yi-1.5-34B-Chat is ranking high on the Chatbot Arena leaderboard! Our model continues to be a top performer, especially across these languages: Spanish: Tie for #1 Japanese: Tie for #2
German: Tie for #3
French: Tie for #3 -
Yi-Large Ranks Among Top Models on MixEval Leaderboard
By
–
We’re thrilled to see Yi-Large ranked among the top models on the MixEval leaderboard! A big thank you to the community for the ongoing support and contributions. Thanks for sharing!
-
Gratitude for ML Community Contributions
By
–
Thank you for highlighting the incredible contributions of the ML community, @Xianbao_QIAN
! -
01.ai Appreciates Feedback on Model Integrations Expansion
By
–
@WolframRvnwlf
, we appreciate your suggestion! We're continuously working on expanding our model integrations, and your feedback is invaluable. Stay tuned for updates ! -
Yi-1.5-34B-Chat Feedback: Planning and Function Calling Success
By
–
Thank you for your feedback! We're thrilled to hear that Yi-1.5-34B-Chat is meeting your needs and excelling in planning and multiple function calling. Stay tuned for more updates and improvements @AtakanTekparmak !
-
Gratitude for Support and Sharing with AI Tool House
By
–
Thank you for your support and for sharing @aitoolhouse !
-
Yi-Large Launch Appreciation and LMSYS Diverse Language Support
By
–
We want to THANK @lmsysorg for ur timely inclusion the day we launched Yi-Large (an unplanned clash w/ GPT-4o, ha). Also very encouraging to see your increasing inclusion of diversified languages (Chinese, French…) and advanced capabilities e.g. launching the latest Hard
-

Yi-Large Ranks #7 Overall and Ties #1 in Chinese on Chatbot Arena
By
–
Thanks 15K+ real user votes on @lmsysorg chatbot arena showing strong results of Yi-Large: #7 in "overall" and tie #1 with @OpenAI GPT-4o in "Chinese." We also found hard prompts, long query, coding useful to challenge LLMs IQs. Try blind test on: https://
arena.lmsys.org