on what possible set of definitions?? and do you think AlphaFold is a conscious being? if how would you draw a line between that and say Claude? it doesn’t hold up to careful inspection.
SAFETY
-
AI capabilities outpace measurement, says Snorkel AI CEO
By
–
“Our ability to do stuff with AI has significantly outpaced our ability to measure capabilities,” says @ajratner, CEO of Snorkel AI.
— Snorkel AI (@SnorkelAI) 31 mai 2026
Alex recently joined the @chain_ofthought podcast with @ConorBronsdon to talk all things frontier AI.
Watch the full conversation covering the… pic.twitter.com/cQZkRoY16i“Our ability to do stuff with AI has significantly outpaced our ability to measure capabilities,” says @ajratner
, CEO of Snorkel AI. Alex recently joined the @chain_ofthought podcast with @ConorBronsdon to talk all things frontier AI. Watch the full conversation covering the -
AI stumble predicted as ‘too big to fail’ with potential carnage
By
–
Anyone remember how I said in January 2025 that AI was going to be stumble and be dubbed “too big to fail”, along with cries for bailouts? If that call was correct – which increasingly seems likely, as the bets escalate to insane heights – the carnage could be immense:
-
Autoreview keeps GPT honest and helps achieve real goals
By
–
gotta use autoreview, that keeps gpt honest and usually helps achieve the real goal.
-
Gary Marcus claims he first noted LLM fabrication before Grok
By
–
all LLMs fabricate things, i was first to point that out in 201, before grok existed
-
Self-recursive improvement, not unsafe release to public
By
–
I think it is more to do with self recursive improvement (I.e. It isn't yet on that path) rather than unsafe to release to the public
-
Hermes allows LLM judgment but constrains irreversible actions
By
–
Hermes handles edge cases by making the skill loop conservative at the boundaries, not by pretending the agent has perfect judgment. The main pattern is: LLM judgment is allowed to propose structure, but irreversible actions are constrained.
-

Claude Opus 4.8 on DeepSWE Bench, 58% Pass@1 and 2nd
By
–
Claude Opus 4.8 has landed on DeepSWE Bench, posting a 58% Pass@1 and taking #2 overall behind GPT-5.5. It continues a broader trend: slightly behind on raw score, but among the most reliable and efficient coding models across recent benchmarks.
-
Confirmed again: LLMs cannot handle the truth, says Marcus
By
–
Why LLMs rarely payoff—and what I have been saying literally for 7 years—confirmed yet again: LLMs can’t handle the truth. (Nor apparently can my critics, who keep saying I am “always wrong”, when I have been saying keeps being confirmed, over and over again.)
-
AI-Powered SecurOS UVSS Detects Explosives Under Vehicles in 3 Seconds
By
–
#AI-Powered SecurOS UVSS Detects Explosives Under Vehicles in Just 3 Seconds
— Ronald van Loon (@Ronald_vanLoon) 30 mai 2026
by @_fluxfeeds
#EmergingTech #Technology #Innovation pic.twitter.com/PaozjezgLE#AI-Powered SecurOS UVSS Detects Explosives Under Vehicles in Just 3 Seconds
by @_fluxfeeds #EmergingTech #Technology #Innovation