8/ Mixture of Memory Experts – proposes an approach to significantly reduce hallucination (10x) by tuning millions of expert adapters (e.g., LoRAs) to learn exact facts and retrieve them from an index at inference time.
Mixture of Memory Experts Reduces LLM Hallucination Tenfold
By
–
