I have been using GPT ImageGen-2 for the past weeks I didn't think that better image-generators would be a big deal but it turns out that there is a quality threshold I didn't expect, where you can now get text, slides, academic papers Look at what it does with my "otter test"!
LLMS
-
Forcing Extended Thinking in Claude Opus 4.7 Image Generation
By
–
This piece thinks that the reason I got a crap pelican riding a bicycle from Opus 4.7 is that it didn't think about it first, I'm trying to figure out if there's a way to force it to think that I've missed
-
Claude Opus 4.7 Extended Thinking API Budget Tokens Availability
By
–
Claude Opus 4.7 with adaptive thinking via the API… am I missing something or is it not possible any more to force it to think? (Prompt hacks like "think step by step" don't count here, I mean the equivalent of budget_tokens or effort: high in previous Claude models)
-

OneVL: One-Step Latent Reasoning and Planning with Vision-Language
By
–
OneVL One-Step Latent Reasoning and Planning with Vision-Language Explanation paper: https://
huggingface.co/papers/2604.18
486
… -
Databricks open source models and datasets release timeline
By
–
Open source models and datasets from databricks when?
-
OpenAI Euphony Tool Mirrors Codex Transcript Viewer Design
By
–
OpenAI's new Euphony tool works almost exactly the same way as my Codex transcript viewer https://t.co/lUQQds4Nvj https://t.co/lI8h72reKy
— Simon Willison (@simonw) 21 avril 2026OpenAI's new Euphony tool works almost exactly the same way as my Codex transcript viewer https://
tools.simonwillison.net/codex-timeline
?url=httpsgist.githubusercontent.comsimonwa9eb5993a2853ec840d26c0e56bde362rawb8c5febdf60d878da84e27c07efdaed159abde4alogs.jsonl#tz=local&q=&type=all&payload=all&role=all&hide=1&truncate=1
… -
API Keys Limitation: Infrastructure Work for Non-Mainline Model Support
By
–
yeah just for API keys right now, it requires some infra work for all the non-mainline models to make support happen for things like agents
-

DeepMind’s Deep Research Max Version Advances Mathematics Capabilities
By
–
New update to DeepMind's Deep Research that takes a leap in capabilities compared to the previous version and incorporates an even more powerful Max version. These are the most "frontier" versions of Gemini, so let's see what fruits it bears in mathematics and research 🙂
-
Stanford AI Chatbots Train Workplace Social Skills Through Role-Play
By
–
Stanford scholars have developed AI chatbots that help people practice and improve essential social skills—like active listening, empathy, conflict resolution, and counseling techniques—through realistic role-play scenarios with personalized feedback. https://
hai.stanford.edu/news/using-llm
s-to-improve-workplace-social-skills
… -
Opus 4.7 Max Shows Regression in Problem Solving Tasks
By
–
Slight regression with Opus 4.7 (Max) though, my guess is that the 'improved instruction following' + lots of thinking about how to solve the problem isn't helping Opus in this case