True story: I stopped thinking about context since GPT 5.3 Codex Single project focused threads with the recent capability of codex to spinoff new threads is goated! Codex continues and goes through compaction but remembers all the important stuff and if not, it’ll look up
LLMS
-

GPT-5.6 ships in three capability tiers, Sol is flagship
By
–
@OpenAI just turned the frontier into a choice.
Instead of one model, GPT-5.6 ships as three capability tiers: Sol — the new flagship. Sets a state of the art on Terminal-Bench 2.1 (complex command-line, multi-step agent work) and is OpenAI's most capable model yet for -

AI outperforms humans in research
By
–
The most interesting result in Anthropic's latest paper isn't the 8x increase in code output. It's this: Claude Mythos Preview suggested a better research direction than humans 64% of the time. We're moving beyond AI that writes code. We're approaching AI that helps decide.
-

@dfintelligence — 2026-06-27
By
–
ATTENTION !!!! Ce serait une très grave erreur de se priver des modèles open source chinois ultra-performants en France. Des réponses orientées ou biaisées, oui, il y en a. Oui, il y en aura. Il y en en a aussi chez Mistral AI, GPT, Anthropic, etc. La probabilité qu'un agent
-

Reuse of Chat, Math, Code categories in open-perfectblend dataset
By
–
Wow, they really reused the three main categories Chat / Math / Code as well https://
huggingface.co/datasets/mlabo
nne/open-perfectblend
… -
Fine-tuning and structured decoding for data extraction
By
–
Yes, for data extraction, I recommend fine-tuning it for your use case (TRL or Unsloth) + using structured decoding (
@dottxtai
) during inference. We'll release something on this topic very soon! -

GPT-5.5-Pro vs Y axis: triangulating spend at $500k/week
By
–

Next in the series of GPT-5.5-Pro vs the Y axis, where we try to triangulate information to add the missing axes. Explanation:
"I anchored the top of the chart at roughly $1M/week because that makes the latest spend about $500k/week, or ~$500 per employee per month for -
Repeating the same context in every chat was unsustainable
By
–
yes, repeating the same context in every chat was unsustainable
-
Access to frontier models cut off, changing the future
By
–
It is absolutely crazy how the last two weeks have changed the entire future. It is unprecedented that access to "frontier" models was cut off,and presumably remains cut off forever. It feels like a watershed moment, as if access to the highest level of human intelligence had