AI Dynamics

Global AI News Aggregator

About

Deakin & Fudan discover Internal Safety Collapse in LLMs

What if your AI suddenly starts generating harmful content while doing a benign task? Researchers from Deakin & Fudan discovered "Internal Safety Collapse" in frontier LLMs. Their TVD framework forces harmful outputs as the only valid completion. Result: 95.3% average safety

→ View original post on X — @jiqizhixin