AI Dynamics

Global AI News Aggregator

About

LayerSkip: Optimizing LLM Inference with Speculative Decoding

LayerSkip is a new paper from @AIatMeta that improves LLM inference by merging speculative decoding and early exit. Authors @m_elhoushi and @AkshatS07 are on alphaXiv to answer questions!

→ View original post on X — @askalphaxiv