Fine-tuning #LLMs isn't just for customization You can also #finetune LLMs to increase throughput using speculative decoding Check out the replay of our recent tech talk to learn how we fine-tuned an LLM using to increase inference by 2x https://
pbase.ai/4bV6rNc.
Fine-tuning LLMs for 2x Faster Inference with Speculative Decoding
By
–
