Can frontier models forecast scientific progress? Mostly no, but here is why. This work looks at 4,760 scientific events across disciplines. Frontier models can identify plausible research directions when given options. They cannot reliably predict whether an advance will land,
LLMS
-

Gemini 3.5 Flash Achieves Pareto Frontier on Vending Bench
By
–
Gemini 3.5 Flash is on the Pareto frontier of cost per intelligence on Vending Bench (a measure of a models ability to run a simulated store)!
-

DeepSeek Sparse Attention Implementation in LLMs Repository
By
–
Added a DeepSeek Sparse Attention (DSA) from-scratch implementation to my LLMs-from-scratch repo thanks to an awesome new reader contrib. With motivation, overview, and GPT-style model reference implementation as standalone example code: https://
github.com/rasbt/LLMs-fro
m-scratch/tree/main/ch04/09_dsa
… -

OpenAI reasoning model reportedly solves decades-old math problem
By
–
80 years. Every top mathematician alive tried this problem. Nobody solved it. An internal OpenAI reasoning model did it in a single attempt. 9 of the world's top mathematicians verified the proof. A Fields Medalist said he'd recommend it for publication "without any
-

A Quick Cheat Sheet to Master Agentic AI
By
–
A Quick Cheat Sheet to Master #AgenticAI
by @genamind #GenAI #LLM #ArtificialIntelligence #ML #MachineLearning #GenerativeAI -
Model tries half context to avoid incompatible downloads
By
–
Yeah and it tries half context if nothing fits at full. Saves you from downloading models that won't work
-

llmfit CLI tool auto-detects hardware and ranks 206 models by VRAM
By
–
Stop guessing which models fit in your VRAM! llmfit is a CLI tool that auto-detects your hardware and ranks 206 models by what actually runs on your system. You download a 70B model and hope it fits. Or you estimate memory requirements across quantization levels and still end
-

AI Glossary: 54 Essential Terms Everyone Should Know
By
–
Your #AI Glossary: 54 Terms Everyone Should Know
by @Imad @jeskillings @cnet Learn more: https://
bit.ly/3RkiSgy #GenerativeAI #LLM #AIAgents #ArtificialIntelligence #MachineLearning -
Using Codex-to-Codex oversight improves AI outputs
By
–
This works because you're adding a layer of oversight between you and the output. When you ask Codex to do something directly, it executes. When you ask Codex to ask Codex, the first agent becomes your project manager. It plans, delegates, and reviews the work before you ever
-
Defending Claude Code’s code as logical and elegant
By
–
Bullshit Nothing ugly about code Claude Code writes, it's logical and elegant
