They found that the probe struggled to extract the bizarro world from the LM, indicating that the original semantics were embedded w/i the LM independently of the probe. This experiment further supported the team’s conclusion that LMs can develop a deeper understanding of
Language Models Develop Deeper Understanding Beyond Probe Extraction
By
–