The most retarded model release yet. What the f*** were they thinking?!??!?
MACHINE LEARNING
-
AI Inference: RDUs, Memory, Parallelism for Token Generation
By
–
What actually happens during AI inference?
— SambaNova (@SambaNovaAI) 9 juin 2026
This video breaks down how RDUs, memory architecture, and multi-level parallelism work together to generate thousands of tokens in parallel across racks.
Built for scalable, real-world AI inference 🦾
Learn more:… pic.twitter.com/YUzEP9hBsrWhat actually happens during AI inference? This video breaks down how RDUs, memory architecture, and multi-level parallelism work together to generate thousands of tokens in parallel across racks. Built for scalable, real-world AI inference Learn more:
-
False positive of the cybernetic classifier, improvement in progress
By
–
Following our exchanges in DM, the problem was a false positive of the cybernetic classifier. The classifier generates many false positives, and we are actively working on improving it. More details here: https://
anthropic.com/news/claude-fa
ble-5-mythos-5
… Apart from that — for the exams of -
Beard disrupts dubbing models; clean-shaven faces recommended
By
–
*disclaimer with my video, my beard is making it hard for most if not all dubbing models out there, hence some shaking and blurring around my mouth in the dubbed version. Yours should be just fine, if you don't have a large facial beard hiding your mouth.
-

XGBoost for Regression, Predictive Modeling, and Time Series Analysis Book Review
By
–
XGBoost for Regression, Predictive Modeling, and Time Series Analysis — Learn how to build, evaluate, & deploy predictive models: http://
amzn.to/4l2YcU9 v/ @PacktDataML —
My review: XGBoost is definitely the focal point and central contribution of this book, along with all -
Larger base model capacity improves training data memorization
By
–
Yes, because the base model has far more capacity than all the previous ones, so it's better at memorizing the training dataset.
-
Allegation that Anthropic nerfed models and Dario sabotaged codebase
By
–
Imagine how long Anthropic have had their models nerfed on purpose Dario masterclass in sabotaging your codebase while you were thinking Claude Code is just working for you lol
-
Insider model access is the new insider trading
By
–
Insider model access is the new insider trading.
-
TIL: AgentsView to Calculate Tokens for Claude Fable 5
By
–
A TIL on using http://agentsview.io to calculate token cost with Claude Fable 5 despite the fact that this model is not yet included in the AgentsView pricing database https://til.simonwillison.net/llms/agentsview-custom-model-price …
-

On-Policy Distillation geometry: fewer weight updates, preserves structure
By
–
“On the Geometry of On-Policy Distillation” OPD is not just SFT mixed with RLVR. It has its own update geometry. This paper shows that OPD updates fewer weights than SFT and preserves pretrained structure better, while staying less constrained than RLVR. The key finding is