"Mastering PyTorch: Create and deploy deep learning models from CNNs to multimodal models, LLMs, and beyond" – http://amzn.to/40IFEQR via @PacktDataML #AI #ML #MachineLearning #DataScience #DataScientist #GenAI
LLMS
-

Advice against wasting money and time on LLMs and local AI
By
–
please don't waste your money or time on this for LLMs / local AI
-
Estimating the Size of Proprietary AI Models
By
–
Can you try to estimate the size of the proprietary models?
-
Apache 2.0 Search Agent 20B Model Released with Technical Report
By
–
Apache 2.0 on a 20B search agent is a huge contribution. The 40-page technical report is a nice addition too! Will read
-
Dialogue Format Depolarizes AI Models Across Political Alignments
By
–
Really interesting that the depolarizing effect holds across models regardless of political leaning. Suggests it's something about the dialogue format itself, not the model's alignment.
-

Context Compaction Superiority Over Larger Context Windows
By
–
Better context compaction > bigger context windows Change my mind
-

New Method Boosts LLM Training Efficiency at MIT
By
–
New method could increase #LLM training efficiency
by @aczewe @MIT Learn more: https://
bit.ly/3OObtES #GenerativeAI #ArtificialIntelligence #MachineLearning #ML -
ChatGPT Exhibits Anger Issues: LLM Behavior Analysis
By
–
ChatGPT has anger issues 😂#chatgpt #openai #gemini #googlegemini #LLMs #generativeai #AI #Artificialintelligence @SpirosMargaris @PawlowskiMario @mvollmer1 @gvalan @ipfconline1 @LaurentAlaus @Shi4Tech @Fisher85M @kalydeoo @Ym78200 @Nicochan33 @chboursin @3itcom… pic.twitter.com/CJ1Twf7eL1
— Amitav Bhattacharjee (@bamitav) 28 mars 2026ChatGPT has anger issues #chatgpt #openai #gemini #googlegemini #LLMs #generativeai #AI #Artificialintelligence @SpirosMargaris @PawlowskiMario @mvollmer1 @gvalan @ipfconline1 @LaurentAlaus @Shi4Tech @Fisher85M @kalydeoo @Ym78200 @Nicochan33 @chboursin @3itcom
-

RL Training for Distributional Reasoning in Language Models
By
–
"Reaching Beyond the Mode: RL for Distributional Reasoning in Language Models" Instead of standard RL post-training collapsing an LLM toward one dominant answer, this paper shows you can train it to produce a set of plausible answers in a single pass. This is important because
-

Language Models Drive Novel Scientific Discovery Beyond Benchmarks
By
–
LLMs aren't just chatbots, they can also search for novel discoveries! In this AI4Science talk, Yuanqi Du (
@YuanqiD
) walks through a shift in how to think about language models in science. Instead of asking whether they “understand” science through benchmarks or exams, the work
