Es el tema del próximo vídeo y es FAS-CI-NAN-TE, pero el crédito va a Anthropic que no sólo ha hecho un increíble trabajo de interpretabilidad, sino también de documentación que os invito a leer si os interesa. https://
transformer-circuits.pub/2024/scaling-m
onosemanticity/index.html
…
Anthropic’s Scaling Monosemanticity: Breakthrough in Transformer Interpretability
By
–
