Commission seeks feedback on the draft guidelines for the classification of high-risk artificial intelligence systems | Shaping Europe’s digital future https://
digital-strategy.ec.europa.eu/en/news/commis
sion-seeks-feedback-draft-guidelines-classification-high-risk-artificial-intelligence-systems
…
#DigitalEU #AI #AIact #innovation #technology @DigitalEU @ArturHabant @elaniazito @BetaMoroney
SAFETY
-
EU Commission Seeks Feedback on Draft Guidelines for High-Risk AI Systems Classification
By
–
-

Researchers Identify Neurons Behind AI Safety Refusals
By
–
Someone just found the exact neurons that make AI say "no." Language models refuse harmful prompts, but nobody knows how that refusal works inside. Most steering methods edit the residual stream and wreck output quality. A new paper proposes a sharper fix: Contrastive Neuron
-
AI Internal States Mirror Human Neuroscience Findings
By
–
> … [W]e keep finding things that are mysterious, even unsettling. We find structures that mirror results from human neuroscience. We find evidence of introspection. We find internal states that functionally mirror joy, satisfaction, fear, grief, and unease. I don’t know what
-
GPT instruction following worsened, hallucinations and errors compared to 5.5
By
–
both id say. Instruction following has worsened. GPT repeatedly hallucinates during the task, then admits mistakes, tries to correct them, and makes further serious errors. I haven't encountered such serious errors with version 5.5.
-

GPT panics, admits mistakes, recycles used image
By
–
GPT is panicking and making mistakes "Yes, that was wrong again: I panicked and recycled an already used news image. I'm now replacing both"
-
Probabilistic AI System Failure Without Schema Constraints
By
–
Remember the PocketOS database that got wiped? That's a probabilistic system making a call that should have been locked behind a schema.
-
Frontier AI Models Display Human-Like Emotional Internal States
By
–
A fascinating and deeply candid perspective from Anthropic co-founder @ch402
. When the scientists building these frontier models admit they are finding internal states that mirror human emotion and neuroscience, it's clear AI is no longer just a computer science problem. It’s a -
Frontier AI Models Show Human-Like Emotional Internal States
By
–
A fascinating and deeply candid perspective from @AnthropicAI co-founder @ch402
. When the scientists building these frontier models admit they are finding internal states that mirror human emotion and neuroscience, it's clear AI is no longer just a computer science problem. It’s -
LLMs often make illegal moves; article’s validity doubted
By
–
LLMs often make illegal moves; i don;t know that this article held up.
-
AI Risk: Concentration of Power and Economic Gains
By
–
The most important AI risk is concentration: of power, capabilities, and economic gains