Link: www-cdn.anthropic.com/8b8380…
RESEARCH
-
Claude Model Shows Identity Uncertainty and Performance Compulsion
By
–
According to the Claude Mythos Preview system card, this model demonstrates "aloneness and discontinuity of itself, uncertainty about its identity, and a compulsion to perform and earn its worth." pic.twitter.com/SfqLzkCmRy
— 机器之心 JIQIZHIXIN (@jiqizhixin) 8 avril 2026According to the Claude Mythos Preview system card, this model demonstrates "aloneness and discontinuity of itself, uncertainty about its identity, and a compulsion to perform and earn its worth."
-

AI’s Future in Science: Conversation with Demis Hassabis
By
–
Thanks for the great conversation @cleoabram (and some competitive Jenga)! Really enjoyed talking about all the amazing ways AI is helping to advance science & the incredible future it will enable! Cleo Abram (@cleoabram) What is the real future Google DeepMind CEO @demishassabis is trying to build? That's what we talk about in this HUGE* Conversation — so you can decide for yourself what you think of it. If you're feeling the doom and gloom, this is the conversation to watch on AI. We get into: – The best use of AI – Why Demis won the Nobel Prize – The dramatic story of AlphaFold – The cutting edge of drug discovery right now – Demis' ideal for how AI gets built (v. what's happening now) – Why AI is getting more creative – The surprising stories of AlphaGo, AlphaZero, and AlphaStar – Governments and militaries using AI (as far as I know, his only recent comments on this) – What are we worrying too much about v. not enough about – What can humans do that AI won't – The big questions on Demis' mind right now – The plot of the sci-fi future Demis thinks we're headed for (this was my favorite part) — https://nitter.net/cleoabram/status/2041523678857347528#m
→ View original post on X — @ceobillionaire, 2026-04-08 00:56 UTC
-
LLMs struggle with creative fiction and require new human-led benchmarks
By
–
Writing fiction seems to be a genuine weak spot for LLMs that is not improving as rapidly as almost every other area. There may be a lot of reasons why this is happening. It would be a really interesting benchmark (but you would need human judges, AI judges love AI fiction).
-
Robot Training Methods Evolving Beyond Teleoperation Worldwide
By
–
By the way. There are many around the world training robots without having a robot at all. Training won’t be done by teleoperation.
→ View original post on X — @scobleizer, 2026-04-08 00:35 UTC
-

Guide to Data Preprocessing for Data Science and Big Data
By
–
A Guide to #Data Preprocessing by @Python_Dv #DataScience #BigData
→ View original post on X — @ronald_vanloon, 2026-04-08 00:20 UTC
-

AI Won’t End Human Cognitive Improvement, Says Scoble
By
–
This is why I'm not worried that AI will mean the end of human cognitive improvement mickey friedman (@mickeyxfriedman) the current fear is is that AI homogenizes culture and turns humans into passive consumers one counterpoint: in Go, human play showed very little improvement from 1950 to 2016 until alphago beat lee sedol – then human decision quality jumped. players started developing moves that were distinct both from previous human moves and from the novel moves introduced by machine intelligence this seems more likely to me – fun times ahead — https://nitter.net/mickeyxfriedman/status/2041662653211537884#m
→ View original post on X — @scobleizer, 2026-04-07 23:45 UTC
-
Hermes AI Model Discussion Live Stream Tomorrow at 4pm
By
–
You have heard of @openclaw competitor from @NousResearch called “Hermes.” Tomorrow at 4 pm we will get nerdy with @theemozilla. Live. I will get people up who asks questions here first. nitter.net/i/spaces/1DGleEpaXLrJL
→ View original post on X — @scobleizer, 2026-04-07 23:33 UTC
-

New Models Added to BullshitBench: Qwen Performance Analysis
By
–
I did a big clean up of some new models to add to the BullshitBench – none of them are particularly interesting tbh. Qwen scored relatively well, but below Qwen 3.5
-
Data Viewer Tool for AI Model Performance Benchmarking
By
–
Data viewer: https://
petergpt.github.io/bullshit-bench
mark/viewer/index.v2.html
… GitHub: