Long Context: (2/6) The long-context capability offered by the Yi model series (Yi-34B-200K and Yi-6B-200K) has gained immense traction from the community. Learn how we extended the Yi base model to 200K long-context through various methods outlined in the paper.
LLMS
-
Yi: Open Foundation Models – New Extensions and Capabilities
By
–
“Yi: Open Foundation Models” by http://
01.AI has arrived. Take a look into the specifics with us and discover the groundbreaking extensions of our Yi model series – which include long context, vision language, depth upscaling, and more. (1/6) -

01.AI Infrastructure Team Showcases FP8 Training at GTC2024
By
–
Hello #GTC2024!! Impressive work done by http://
01.AI amazing infra team, presented as a @nvidia official best practice for end-to-end FP8 training and inference. Our larger model is under way -
Comparative Analysis of ChatGPT and Gemini AI Capabilities
By
–
AI showdown: ChatGPT vs Gemini. Deciding factors in AI utility. • Compare features
• Evaluate performance
• Understand use-cases Read more: https://
buff.ly/49WmCJG https://
buff.ly/49WmCJG -
Sam Altman Discusses GPT-5 Capabilities and AGI Vision
By
–
On a Lex Fridman podcast released today, Sam Altman spoke on GPT-5, Sora, his vision for AGI, and more.
— Rowan Cheung (@rowancheung) 19 mars 2024
When asked about GPT-4, Altman said it “sucks“ and that the leap in capabilities for GPT-5 will be similar to GPT-3’s jump to GPT-4. pic.twitter.com/Of1Vh5l0NzOn a Lex Fridman podcast released today, Sam Altman spoke on GPT-5, Sora, his vision for AGI, and more. When asked about GPT-4, Altman said it “sucks“ and that the leap in capabilities for GPT-5 will be similar to GPT-3’s jump to GPT-4.
-
Sam Altman discusses GPT-5 future amid major AI developments
By
–
AI NEWS: Sam Altman just spoke about the future of AI and GPT-5. Plus, more developments from Apple, Google Gemini, NVIDIA GTC, Mercedes-Benz, and Stability AI. Here's everything going on in AI right now:
-
Scaling Trends: Establishing Testing Standards for Next-Gen AI Models
By
–
Also, sets up a good standard for testing when GPT-5, Gemini 2.0, etc. come out. We need to understand scaling trends to see what progress is actually being made, as the benchmarks out there are not very useful.
-
Testing GPT-4 Papers Against Gemini 1.5 and Claude 3
By
–
I want to see key GPT-4 papers re-tested with Gemini 1.5 and Claude 3 to see what generalizes across GPT-4 class LLMs. At a minimum, the papers on hallucination rates, Theory of Mind & Chain of Thought; as well as papers on performance on medical, legal & psychological questions
-
NVIDIA Blackwell Platform Unleashes Real-Time Generative AI
By
–
Delivering a massive upgrade to the world’s #AI infrastructure, our CEO Jensen Huang introduced the NVIDIA Blackwell platform to unleash real-time generative AI on trillion-parameter LLMs at today's #GTC24 keynote. Read more about our announcements. https://
nvda.ws/48Z5DoO
