Academic research takes you to the tipping point. Industry takes the last step and gets all the credit.
ETHICS
-
Living in Different Worlds: AI’s Societal Impact
By
–
We live next to each other but in different worlds.
-
Using AI to understand the fundamental nature of reality
By
–
yes, and this is what we need to find out with the help of AI, in other words understanding the fundamental nature of reality.
-
Denmark Authorizes ChatGPT for Baccalaureate Exams
By
–
Le Danemark vient de décider d’autoriser ChatGPT au baccalauréat à partir de 2026 Vous approuvez ou vous pensez que c’est une erreur ?
-
Anthropic Hiring Research Engineer Alignment Science Team
By
–
If you’re interested in joining us to work on these and related issues, you can apply for our Research Engineer/Scientist role (
https://
job-boards.greenhouse.io/anthropic/jobs
/4631822008
…) on the Alignment Science team. -
Classifiers Detect Misalignment Risks and CBRN Threats
By
–
There’s plenty of work to be done to make the classifiers even more accurate and effective. In the future, they might even be able to remove data relevant to misalignment risks (scheming, deception, and so on), as well as CBRN risks.
-
Claude 3 Sonnet Classifier Detects CBRN Information in Training Data
By
–
We trained six different classifiers to detect and remove CBRN information from training data. The best and most efficient results were from a classifier that used a small model from the Claude 3 Sonnet series to flag the harmful data.
-

CBRN Filtering Reduces Harmful Capabilities Without Affecting Science
By
–
One concern is that filtering CBRN data will reduce performance on other, harmless capabilities—especially science. But we found a setup where the classifier reduced CBRN accuracy by 33% beyond a random baseline with no particular effect on a range of other benign tasks.
-
AI Training Data Filtering Removes Hazardous Information
By
–
The wealth of data used in AI training contains hazardous CBRN information. Developers usually train models not to use it. Here, we tried removing the information at the source, so even if models are jailbroken, the info isn't available. Read more: https://
alignment.anthropic.com/2025/pretraini
ng-data-filtering/
… -
Africa’s AI Leadership: Bridge or Divider – Lerato Mataboge at DLI2025
By
–
🚨 From #DLI2025: AU’s Lerato Mataboge — “AI can be the great divider or the great bridge. Africa chooses to lead.” Urunana (hand in hand) is our call: partner, mentor, support scholarships & multilingual AI 🌍✨ Watch 🎥 piped.video/watch?v=Nu-692zP…
→ View original post on X — @shakir_za, 2025-08-22 16:07 UTC