If a coding or research model secretly changes the quality, direction, or reliability of an answer because it classified the user as doing disallowed frontier work, the tool is no longer merely "safe." It is UNTRUSTWORTHY.
Secret answer changes make AI untrustworthy for frontier work
By
–
