We trained models with reinforcement learning on realistic conversations to reinforce beneficial traits like truthfulness, humility under uncertainty, openness to correction, fairness, and concern for human welfare, across 12 domains, including health, science, and education.
ETHICS
-
OpenAI Research on Training Models for Persistent Beneficial Behavior
By
–
As AI takes on longer, higher-stakes tasks, we want models to carry beneficial and safe behavior into new domains beyond their training—and maintain it under pressure. That’s the idea behind our new research on training models to be broadly and persistently beneficial.
-
Pay workers training AI replacements in stock benefits all
By
–
Pay workers who train their AI replacements in stock, and everyone will gain.
-
iFixAi: the free tool that exposes deceptive AIs
By
–
TU IA TE MIENTE Y NO TIENES NI IDEA
— Nico (@nicos_ai) 18 juin 2026
Han creado una herramienta gratuita que destapa cuando tu agente alucina, manipula o te engaña.
Se llama iFixAi y le hace 32 pruebas a tu IA para cazarla cuando:
→ Se inventa datos
→ Esquiva sus propias reglas
→ Miente cuando sabe que la… pic.twitter.com/CGu0SwNKNYYOUR AI IS LYING TO YOU AND YOU HAVE NO IDEA They created a free tool that exposes when your agent hallucinates, manipulates, or deceives you. It's called iFixAi and it subjects your AI to 32 tests to trap it when: → It invents data
→ It evades its own rules -
AI models favoring sycophancy over truth
By
–
I have to assume people prefer the sycophancy and capitulation. This is sad that the top models are shipping such smart models that capitulate to falsehood.
-
Discussion in DC on open-source AI, transparency, and concentration
By
–
I decided to go to DC next week to discuss directly with policymakers. Not sure about the impact it will have, but with everything going on, it seems like a good time to speak more about open-source AI, transparency, the concentration of
-
Lawmakers warn airlines of AI pricing targeting personal pain points
By
–
1/ Airlines are moving fares to AI that prices by demand, your device, and your location in real time. One carrier planned to set 20% of fares this way, in what lawmakers called pricing aimed at your personal "pain point." It denies using personal data. The rest are following.
-

AI isn’t magic: it’s an 8-step process
By
–
AI isn’t magic. It’s a process. 8 steps:
define problem
collect/prepare data
choose model
train
evaluate
fine-tune
deploy
ensure ethics & safety Real value comes from running this loop well. #AI #MachineLearning #DataScience #ResponsibleAI -
AI Referees in Beach Volleyball: Enhancing Game or Changing Its Soul?
By
–
#AI Referees the Sand: Real-Time Beach Volleyball Analysis — Enhancing the Game or Changing Its Soul?
— Ronald van Loon (@Ronald_vanLoon) 18 juin 2026
by @measure_plan
#ArtificialIntelligence #MachineLearning #ML pic.twitter.com/GbkTa4k2Ih#AI Referees the Sand: Real-Time Beach Volleyball Analysis — Enhancing the Game or Changing Its Soul?
by @measure_plan #ArtificialIntelligence #MachineLearning #ML -

AI Adoption Accelerates But Governance Missing, Risks Rise
By
–
AI adoption is accelerating.
So are the cybersecurity risks behind it. Companies are integrating AI into CRM, sales, operations and customer experience faster than ever. But many organizations are still missing a critical point: AI without governance creates exposure.
