Unless and until agents really do work at expert level, the benefits of AI use are going to be contingent on the skills of the AI user, the jagged abilities of the AI you use, the process into which you integrate use, the experience you have with the AI system & the task itself
@emollick
-

GitHub Copilot Study: AI Transforms Developer Work and Earnings
By
–
This large study of 187k developers using GitHub Copilot finds AI transforms nature of coding. Coders focus: more coding & less management. They need to coordinate less, working with fewer people They experiment more with new languages, which would increase earnings $1,683/year
-

AI Tools Adoption: Learning Curve and Process Adjustments
By
–
I think one factor driving varied results is that AI tools are actually not that easy to use right away and require a learning curve of some hours (see @simonw
) Another is that they require adjustments in process A third is that AI has different uses depending on user expertise -

Large-scale coder field experiments clash with AI workforce trends
By
–
This clashes with some large-scale field experiments of coders working at companies, but worth noting.
-

AI Image Generation 3D Game Engine Odyssey Reaches New Capability
By
–
AI image generation 3D "game engines" seem to have reached some critical ability threshold recently (not there yet of course, but much improved over a couple months ago). Here is another one, Odyssey, that lets you wander around virtual worlds where each frame is AI generated. pic.twitter.com/Lv0Zo5qfiJ
— Ethan Mollick (@emollick) 10 juillet 2025AI image generation 3D "game engines" seem to have reached some critical ability threshold recently (not there yet of course, but much improved over a couple months ago). Here is another one, Odyssey, that lets you wander around virtual worlds where each frame is AI generated.
-
O3 vs Grok: Sophistication Projection and Competitive Pivot Analysis
By
–
As a business school professor, I think they both did fine but had gaps (o3 did a more sophisticated projection, Grok did a better pivot in the face of competition). Can't penalize Grok for no image creation, but it is generally is less tool-using & "agentic" than o3 currently.
-

Grok 4 Documentation Release Status and Availability
By
–
Is there any documentation for Grok 4 anywhere yet? The xAI website last mentions the Grok 3 beta, no new prompts on the Github, etc.
-
Grok 4 Analysis: Hidden CoT, Web Search Integration
By
–
A few quick observations on Grok 4:
1) Hidden CoT with very little information in the reasoning trace
2) Uses web search a lot (not just searching X)
3) Have not seen it use code to run calculations or solve non-coding problems yet, generally less aggressive about tools than o3 -
Company AI API Lacks Transparency After Model Incident
By
–
Hard to imagine a lot of excitement inside companies for using the API at this stage, given the lack of transparency – I am surprised they didn't take advantage of all the eyeballs to offer some sort of post mortem on yesterday. They pulled their own model down as a result.
-
xAI Grok 4 RonnaFLOP model performance benchmarks scaling
By
–
I suspect the next few weeks after Grok 4 follows the same pattern as Grok 3 xAI beats everyone to market with the first RonnaFLOP model. The benchmarks show the 10-20% improvement the scaling law suggests. In the coming months, the other labs release their RonnaFLOPs, catch up.