1/ Scale is announcing our latest SEAL Leaderboard on Adversarial Robustness! Red team-generated prompts Focused on universal harm scenarios Transparent eval methods SEAL evals are private (not overfit), expert evals that refresh periodically http://
scale.com/leaderboard
Scale SEAL Leaderboard Evaluates AI Adversarial Robustness
By
–
