NEW RESEARCH: Approximating Language Model Training Data from Weights ever wonder how much information is available in an open-weights model? DeepSeek R1 weights are 1.2 TB… what can we learn from all those bits? our method reverses LLM finetuning to recover data:
Reverse Engineering Training Data From Language Model Weights
By
–
