5/ They are open-sourcing the Lite series in two sizes, an 8B dense model and an A3B mixture-of-experts version. The product has some exciting new updates. An 8-step distilled LoRA now open-sourced: inference cut from 23s to 2s on H100 (100 NFE -> 8 NFE).
ComfyUI is now
LLMS
-
Open-sourcing Lite series with 8B and MoE models plus fast distilled LoRA
By
–
-

Mix-Quant: Quantized Prefilling and Decoding for Agentic LLMs
By
–
Mix-Quant Quantized Prefilling, Precise Decoding for Agentic LLMs
-

Auth Proxy: Controlling Agent Behavior Boundaries
By
–
Introducing the sandbox Auth Proxy: A way to control the boundary between agent-generated behavior and the rest of the world. An explainer from @hwchase17
-
Token Economics: LLM Utility, Demand, Supply, Monetization
By
–
→Token utility: the model, context length, and level of interactivity required
→ Token demand: the volume of tokens required
→ Token supply: the optimal infrastructure given the token utility and demand
→ Token monetization: cost-based and value-based pricing to maximize -

LLM Hallucinations and Visual Misinterpretation Research
By
–
Yowza! Similar in some ways to the Stanford paper recently on LLMs hallucinating responses to images they never saw.
-

Tapping Termius opens sites with Claude Code on VPS instantly
By
–
Oh ok so it's super easy for me I just tap Termius and each of my sites is open with Claude Code open on the VPS instantly
-
Google accused of lazy benchmaxxing of models
By
–
Nah, thankfully Google is lazy and just benchmaxxing their models at this point
-
Gentle reminder: Anthropic and OpenAI are not your friends
By
–
Gentle reminder Anthropic & OpenAI are not your friends Dario and Sam Altman are mercenaries who would love for no one else but themselves to control that blackbox we call AI Don’t let PR campaigns & limit resets make you forget that they have every intention of rugpulling you
-

Evaluating Memory in Long-Horizon AI Agent Systems
By
–
LongMINT Evaluating Memory under Multi-Target Interference in Long-Horizon Agent Systems
-

Composer 2.5 scores 62, price difference makes ‘slightly better’ not worth 60x
By
–

Composer 2.5 scores 62 on the Artificial Analysis Coding Agent Index. The two models above it score 65 and 66. The price difference: $0.07 per task vs. $4–5. At some point "slightly better" stops being worth "60x more expensive," and most engineering teams crossed that point a