Quick update to GPT-5.5 / Spud: AI labs are shifting back toward making models smarter during pretraining rather than leaning on reasoning (test-time compute) to boost performance. OpenAI's Spud and Anthropic's Mythos both appear to reflect this trend, getting better answers
LLMS
-
Team collaboration optimizing AI models in late night sessions
By
–
Late night Slack threads with the team when we’re all cooking with our models are the best threads!
-

ChatGPT Image 2.0: Man returns canvas print of himself being photographed at HomeGoods.
By
–
ChatGPT Images 2.0 (Pro) generates an amateur photo of a man returning a decorative canvas print to HomeGoods because the print is actually just the same candid photo of him being taken at that moment and that doesn’t make any sense.
-
Knowledge of Wikipedia URLs for numbers without searching
By
–
It surely knows the Wikipedia URLs for numbers without searching. I’m not sure it searched anything.
-

OpenAI Releases Free Healthcare ChatGPT-5.4 for Clinicians
By
–

Interesting, OpenAI just released a free healthcare version of ChatGPT-5.4 for clinicians that beat specialty-matched physicians with unlimited time + web access on a benchmark of real & hard clinical tasks. Caveat: the benchmark was designed by OpenAI, though it is fully open.
-

Moonshot Releases Kimi K2.6 Open-Weight Model
By
–
Kimi K2.6, the new state-of-the-art open-weight model from Moonshot, is now available for Pro and Max subscribers.
-
GPT 5.5 Release Expected Tomorrow
By
–
GPT 5.5 tomorrow would be the best damn birthday gift I could ever ask for
-

DataFlex: Dynamic Data Optimization for LLM Training
By
–
What if you could supercharge LLM training by dynamically optimizing the data, not just the model? Researchers from Peking University, Shanghai AI Lab, & the LLaMA-Factory team present DataFlex. It's a unified framework that smartly selects, mixes, and re-weights training data
-

OpenAI Launches ChatGPT for Clinicians and HealthBench Professional
By
–
Today @OpenAI introduced ChatGPT for Clinicians, provided free for credentialed HCPs, and HealthBench Professional for benchmarking LLM medical task performance (Figure) https://
openai.com/index/making-c
hatgpt-better-for-clinicians/
… https://
cdn.openai.com/dd128428-0184-
4e25-b155-3a7686c7d744/HealthBench-Professional.pdf
… -
Rewarding LLM Uncertainty Reduces Hallucinations
By
–
New @Nature To reduce LLM hallucinations they should be rewarded for admitting uncertainty
