GitHub: https://
github.com/thu-pacman/chi
tu
…
By the way, Chitu (赤兔) is the name of a great horse in ancient China.
@jiqizhixin
-
Chitu: Ancient Chinese Horse-Named AI Project on GitHub
By
–
-
Development Team Optimizes GeMM and MoE for FP8 Processing
By
–
A member of the development team told us: "We've optimized a series of key operators, such as GeMM and MoE, at the instruction level to achieve native processing capabilities for FP8 data."
-

Chitu Matches vLLM Performance on H20 DeepSeek-R1
By
–
On an H20 (96GB) cluster, Chitu performs comparably to vLLM when deploying DeepSeek-R1-671B.
-

DeepSeek-R1-671B Outperforms vLLM on A800 Cluster
By
–
And it outperforms vLLM when deploying DeepSeek-R1-671B on an A800 (40GB) cluster.
-
Chitu: High-Performance Open-Source LLM Inference Framework
By
–
Yet another open-source project from China: Chitu! A high-performance inference framework for LLMs, designed for efficiency, flexibility, and availability. Chitu supports various mainstream models, including DeepSeek, the LLaMA series, Mixtral, and more.
-

Meta researchers replace normalization with Dynamic Tanh in Transformers
By
–
Imagine a Transformer model without normalization. This is exactly what's proposed in a new paper from Meta, NYU, MIT, and Princeton. The authors found that normalization layers can be replaced with something called Dynamic Tanh (DyT). It looks like this: DyT(x)=γ ∗ tanh(αx)+β.
-
Chain of Thought Enhances AI Image Generation Step by Step
By
–
Can We Generate Images with CoT? Let’s Verify and Reinforce Image Generation Step by Step https://
arxiv.org/pdf/2501.13926 -
ByteDance Introduces OmniHuman-1 for Conditioned Human Animation
By
–
🔥🔥🔥ByteDance introduces OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Modelshttps://t.co/nGWC8HyoLEhttps://t.co/CSzIVEFT8i pic.twitter.com/BWwMelPbfq
— 机器之心 JIQIZHIXIN (@jiqizhixin) 7 février 2025ByteDance introduces OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models https://
omnihuman-lab.github.io https://
arxiv.org/abs/2502.01061 -
LIMO: Less Data Improves Reasoning Performance
By
–
LIMO: Less is More for Reasoning With only 817 examples, it outperforms previous models, proving less data can enhance reasoning. https://
arxiv.org/pdf/2502.03387 https://
github.com/GAIR-NLP/LIMO https://
huggingface.co/datasets/GAIR/
LIMO
… -

ByteDance Releases Doubao-1.5-Pro AI Model
By
–
ByteDance Doubao-1.5-pro is here https://
team.doubao.com/en/special/dou
bao_1_5_pro
…