GLM 5.2 appears to be a real turning point for open source, challenging proprietary AI models.
LLMS
-

The bible for running LLMs locally now free online
By
–
DROP EVERYTHING The bible for running LLMs locally is now available online to read for free Covers what to use on – Laptop / edge / odd hardware
– Mac-first workflows
– Single RTX GPUs
– 2-4+ NVIDIA / CUDA GPUs
– General production serving
– Long-context / MoE / routing
– -
AI finds errors and updates grad school paper with new data
By
–
The interaction between AI & past scholarly work is going to get weird. Here I gave GPT-5.5 Pro a copy of my first published paper from grad school & asked it to find errors and update it. It found new data, analyzed it, created reproducible files, extended the key argument…
-
AI industry on X discusses GLM and Hermes, lacks diverse voices
By
–
The AI industry on X loves to talk about technology. This weekend’s hot tech is GLM, an AI Open Source model. Previous weekends saw a ton of discussion about how @NousResearch
’ Hermes was better than @openclaw
. One problem with X is that we don't hear enough from people -
Stop hardware cost to token calculations, models improve, prices rise
By
–
Can we stop doing hardware cost to token generation calculations on the timeline please? If you haven't noticed, models keep getting better & more efficient, and hardware prices keep going up
-
ColBERT outperforms on CPU with low latency for embeddings
By
–
I'm talking about individual descriptions used for embeddings. It doesn't need to be particularly long for late interaction to perform better. The tradeoff really depends on the use case. In this case, even on a cheap CPU, the latency is so low that ColBERT just works better!
-

Comparison of LLM, RAG, AIAgent, and MCP
By
–
#LLM vs. RAG vs. #AIAgent vs. MCP
by @Python_Dv #GenerativeAI #ArtificialIntelligence #MachineLearning #ML -

Meta limits internal use of AI amid exploding token costs
By
–
No more tokenmaxxing at Meta Meta is preparing to limit internal use of AI after employee token consumption surged so much that the company now expects internal AI costs alone to reach billions of dollars in
-

Luke Alonso uploaded NVFP4 of GLM 5.2, 467GB on 4 DGX Sparks
By
–
Luke Alonso has uploaded an NVFP4 of GLM 5.2 467GB, would fit on 4x DGX Sparks (~$20k)