Iterative improvement: Freeze the model + environment while you tweak the rubric, run agreement evals + collect expert reviews to score, make number go up
GENERATIVE AI
-

Qwen3 Model Matches or Exceeds Larger Family Models
By
–
A la par o superando incluso a otros modelos más grandes y pesados de la familia Qwen3
-

Qwen3-Next-80B-A3B: Efficient MoE Model with Superior Performance
By
–
Qwen3-Next-80B-A3B is out 80B params, but only 3B activated per token → 10x cheaper training, 10x faster inference than Qwen3-32B.(esp. @ 32K+ context!) Qwen3-Next-80B-A3B-Instruct approaches our 235B flagship. Qwen3-Next-80B-A3B-Thinking outperforms
-
FFMPEG binary integration in Claude container for enhanced capabilities
By
–
Install the FFMPEG binary in the Claude container for us and we can have even more fun with that kind of thing!
-

RAG vs Agentic RAG: Key Differences Explained
By
–
What’s the difference between #RAG and #AgenticRAG? http://
bit.ly/4pkl7Nc @bytebytego #AI #AIAgents #AgenticAI @mvollmer1 @Kevin_ODonovan @Shi4Tech @enilev @CatherineAdenle @FmFrancoise @AkwyZ @jblefevre60 @EvanKirstel @BetaMoroney @gvalan @CurieuxExplorer @Fabriziobustama -
Developer products and open source strategy for personal AI
By
–
Is he expecting to build any developer aimed products? What would be his open source strategy? Or is personal super intelligence all about chatgpt in WhatsApp?
-
Western AI Censoring Causes More Production Issues Than Chinese Models
By
–
For a production app the Western AI models censoring is much more annoying than the Chinese censoring In 0.0001% of cases someone will try to generate Tiananmen Square But in 80% of cases they generate a normal video which gets false flagged as NSFW with Veo 3 for example https://
t.co/RfK54OMvSC -
The Algorithm as Dragon: AI Mythology and Cultural Narrative
By
–
the algo is the dragon in this context pic.twitter.com/p9O6mU1RNP
— Ahmad (@TheAhmadOsman) 11 septembre 2025the algo is the dragon in this context

