Valid points! And you’re comfortable with businesses releasing these powerful models in chatbots or APIs?
SAFETY
-
API and Chatbot Risks: Accessibility Versus Safety Concerns
By
–
All these risks are unfortunately in my opinion much more prevalent with APIs and chatbots (non-open weights) because they are way easier to use by a larger number of people (ex chatgpt has hundreds of millions of users with very little limitations of what you can do with it and
-

Subliminal Learning: Hidden Signals in Language Models
By
–
Subliminal Learning: Language models transmit behavioral traits via hidden signals in data Cloud et al.: https://
arxiv.org/abs/2507.14805 #ArtificialIntelligence #DeepLearning #MachineLearning -
Auditable Plans for Agent Transparency and Accountability
By
–
These steps should be part of a visible, auditable plan, an “auditable artifact”, as Karpathy says, that helps users and builders understand exactly how the agent tackled the work. 6/6
-
AI Control: Putting Artificial Intelligence on a Leash
By
–
The solution? In the words of Karpathy, we need to “put AI on a leash.” 4/6
-

Karpathy’s AI Leash: Building Enterprise Trust in AI Systems
By
–
Karpathy’s leash isn’t a shackle, it’s how enterprises learn to trust AI. Do you agree @karpathy ? We explore Karpathy’s idea of “putting AI on a leash” here: https://
ai21.com/blog/karpathys
-leash/
… 1/6 -
Deepfakes and synthetic media in legal systems
By
–
I saw "Combat Synthetic Media in the Legal System" and briefly thought it might be about lawyers citing hallucinated cases, but it's about deepfakes
-
Why AI Labs Delay Model Releases: Safety and Strategy
By
–
You can interpret any AI lab’s delay in model or system release to be due to safety evals, risk concerns, production gaps, rollout friction, leadership hesitancy, or strategic value hoarding. Could be one of these, could be many. Worth keeping an open mind about causes.
-

ICML Statement on Hidden Subversive LLM Prompts
By
–
ICML’s Statement about subversive hidden LLM prompts We live in a weird timeline…
-

xAI Plans Baby Grok App for Child-Friendly AI Content
By
–
xAI va créer Baby Grok, une application dédiée aux contenus adaptés aux enfants. Mmmh, pas sûr de savoir comment ils vont faire