We used weak supervision to programmatically curate instruction tuning data for open-source LLMs like Llama 2 and RedPajama, enabling more granular error analysis and higher quality—without an army of manual annotators. Links to data and models on the blog!
Weak Supervision for Open-Source LLM Instruction Tuning Data Curation
By
–