Reinforcement learning is time travel for rewards.
RESEARCH
-
ULMFiT LSTM Evolution Language Model Fine-tuning History
By
–
Right, and indeed ULMFiT was an LSTM. But also, that earlier work wasn't using language modeling on a general purpose corpus as a self-supervised task then fine-tuning that in two more steps for downstream tasks like today's LLMs. (It wouldn't have been feasible on that h/ware)
-

Heart-Brain Axis: Atrial Fibrillation Impact on Brain Glymphatic Flow
By
–
Heart-Brain Axis
Atrial fibrillation leads to reduced brain glymphatic flow, decreased washout of waste metabolites. https://
academic.oup.com/eurheartj/arti
cle/46/18/1733/8029578?login=false
…
Today #ACC26 @JAMACardio Serum Neurofilament (sNfl), a marker for brain cell injury, is associated with adverse CV events and mortality in -
Research-driven AI company focused on practical applications
By
–
Fully opensource stack, and I trust their mission than I trust Intel (which NVIDIA is in an investor in btw) Do people not realize that we need healthy competition for local inference? I am hoping to actively help mature that stack myself btw
-
Codex Security Research Preview Now Available for Testing
By
–
Seeing a lot of interest here! If you want to try Codex Security, you can read more about how it works and early findings on our blog: https://
openai.com/index/codex-se
curity-now-in-research-preview/
… And here’s how to get started: https://
developers.openai.com/codex/security
/setup
… -
Intelligence and Skill Are Not Equivalent Concepts
By
–
Bad analogy because you're equating intelligence and skill (being good at chess or Go), but intelligence is emphatically *not* skill
-

Serial Experiments Lain 1998 – Machine Learning References
By
–

Serial Experiments Lain (1998) nitter.net/hardmaru/status/828047… hardmaru (@hardmaru) Machine Learning Enthusiast — https://nitter.net/hardmaru/status/828047097366667264#m [Translated from EN to English]
-

dLLM: Simple Diffusion Language Modeling Framework Overview
By
–
dLLM: Simple Diffusion Language Modeling buff.ly/9kxHBtV Although diffusion language models (DLMs) are evolving quickly, many recent models converge on a set of shared components. #AI #MachineLearning #DeepLearning #LLMs #DataScience
→ View original post on X — @miketamir, 2026-03-29 16:49 UTC
-
Towards end-to-end automation of AI research
By
–
Towards end-to-end automation of AI research
https://www.nature.com/articles/s41586-026-10265-5 [Translated from EN to English]→ View original post on X — @sakanaailabs, 2026-03-29 16:35 UTC
-

Meta tests multiple Avocado AI variants internally
By
–
BREAKING : Meta is testing loads of Avocado variants internally, including multiple release candidates, Avocado-mango agent, Avocado 9B, and more. Avocado Think Hard performs quite well and, as reported earlier, is comparable to Gemini 3 level models. All this in parallel to