ByteDance is reportedly building its own inference chip modeled on Groq's LPU, the same architecture Nvidia paid roughly $20B to license in December. The LPU keeps the model in on-chip SRAM and skips high-bandwidth memory. HBM is the component the US restricts most tightly for
ByteDance Reportedly Building AI Inference Chip Modeled on Groq’s LPU Architecture
By
–
