Very cool careful paper that shows that management consulting has a direct, positive (but small) effect on businesses in Belgium. Companies spend 3% of their revenue on engagements, and get "positive effects on labor productivity of 3.6% over five years" and increases in wages.
@emollick
-

Open Source AI Models Show Inconsistent Performance vs DeepSeek
By
–
These new open source models (GLM, Kimi) continue to be odd. Great stats, some solid performances, but also fail tests that DeepSeek & smaller closed models have beaten for months.
-
AI Adoption in Organizations: Integration Challenges and Agent Lessons
By
–
Right now, AI adoption in organizations is constrained by the need to figure out how to integrate AI with the complex & often poorly-understood processes inside companies But ChatGPT agent suggests that The Bitter Lesson of AI may come for real work, too. https://
open.substack.com/pub/oneusefult
hing/p/the-bitter-lesson-versus-the-garbage?utm_source=app-post-stats-page&r=i5f7&utm_medium=ios
… -
Kimi K2 Model Exhibits Significant Hallucination Issues
By
–
Kimi K2 really hallucinates a lot in my limited testing so far, and is very happy to make up new details if it judges that it would improve the punch of a paragraph.
-

LLM Arena Models Proliferation: Concerns About Score Gaming
By
–
Suddenly there are tons more weird LLM arena models – cuttlefish, kraken, etc. I just hope we are not going to see a repeat of the Llama 4 incident, where different versions of the same model are being tuned to max out the arena score
-

Limited Research on AI Impact Across Key Professions
By
–
There are far too few careful studies of the progress of AI in key professions and fields that may be most impacted Example: there are only a couple of good controlled studies on lawyers working with AI, the most recent used (now obsolete) o1-preview & even that had big effects.
-
Best AI Models Capabilities Six Months Ago
By
–
The best models could do 6 months ago. pic.twitter.com/Kg5iAczkaS
— Ethan Mollick (@emollick) 27 juillet 2025The best models could do 6 months ago.
-
Generative Starship Control Panel Code with p5.js
By
–
Just the two prompts, exactly as depicted: create something I can paste into p5js that will startle me with its cleverness in creating something that invokes the control panel of a starship in the distant future and make it better
-
Prompting AI Systems to Avoid Reinforcement Learning from Human Feedback
By
–
Gotta ask it to make stuff for you or you get RLHF'ed
-
Testing AI Models Through LMArena: Summit Model Access
By
–
This is through LMArena, where you are given random models to test. You will likely get a chance to use "Summit" fairly often (it came up three times in my six attempts):
