Whatever people build around the model today becomes training data for the next one, that loop just keeps tightening to, TBH, a worrying point haha. Quite excited/scared to see where this "slopacopalypse" will lead us.
@whats_ai
-
Semantic Correctness Over Exact Match for Agent Evaluation
By
–
Semantic correctness over exact match is the right call for agent work, exact match penalizes paraphrases the downstream agent doesn't even care about.
-
Agent Markets Concentrate on Model Quality Over Routing Strategy
By
–
The "higher quality model wins" finding is spicy, suggests agent markets will concentrate around model quality more than routing cleverness. In the end, we have to have subscriptions to all providers lol
-
Compute Wars Intensifying: The Race for Computing Supremacy
By
–
Compute wars intensifying. What a time to be alive haha.
-
Open Source Code Execution Chat ML Innovation
By
–
Will play around! Open source execution over chat is what ML work has been waiting on.
-

ChatGPT User Demographics Shift Toward Gender Parity
By
–
Really cool data from OpenAI: ChatGPT went from 83% typically masculine first-name users in Jan 2023 to roughly 50/50 today. 34 points in 40 months! (Redesigned OpenAI's chart because the original was pretty ugly, data is theirs.)
-

Recursive Self-Improvement Loop: From AutoResearch to AGI
By
–
Recursive self-improvement is not AGI hype. And it’s not just prompt tuning either. Karpathy’s autoresearch ran 700 experiments on a single GPU, improved its own training code, and kept what worked. This loop is already here. I break down how it works, where it fails, and how
-
Six Years of Weekly AI Engineering Content Posts
By
–
I’ve been posting every week for 6 years, and I’m going all in on AI engineering content. Subscribing is free and it helps a lot. Thank you for being part of the journey:
-
Tool Use and Self-Check as Core AI Workflow Components
By
–
Excited to compare in my workflows. Tool use + self-check is basically the whole game now.