Files are long term memory. Context windows are working memory. That's all you need.
LLMS
-
ConceptMoE: Adaptive Token-to-Concept Compression for Computing
By
–
ConceptMoE Adaptive Token-to-Concept Compression for Implicit Compute Allocation
-
Scaling Embeddings Outperforms Scaling Experts in Language Models
By
–
Scaling Embeddings Outperforms Scaling Experts in Language Models
-
Plain English as the Natural Interface for Data Tools
By
–
Plain English to insights is the interface data tools should've had all along.
-
Technology for AGI: beyond content creation
By
–
I see, but beware, this technology is designed to eventually train AGI. For content creation it's overkill. The ultimate engineering goal here is not to create content on demand (even if that will be possible) but rather to see the limits of models for
-
Understanding What Large Language Models Actually Are
By
–
A lot of folks without much background in building LLMs misunderstand what an LLM actually is.
-
LLMs as Part of Intelligent Systems
By
–
Yes very exciting direction – an LLM can be part of a *system* that could be very helpful.
-
Scobleizer on world models vs LLMs and video training
By
–
I've been yelling and screaming that the future is world models after Oliver Cameron showed me how. World models are not LLMs (large language models). World models are trained mostly with video data or other kinds of data that isn't really text.
-

LLM Engineer’s Handbook for Training and Production
By
–
LLM Engineer's Handbook — Master the art of engineering Large Language Models #LLMs from concept to production: http://
amzn.to/4dUQrv6 v/ @PacktDataML Implement robust data pipelines and manage LLM training cycles Create your own LLM and refine with the help of hands-on -
Reasoning Models: The Next AI Leap Toward Inspectable Thinking
By
–
Why Reasoning Models Are The Next Leap In AI #Reasoning models mark a critical step forward in #AI, moving systems from fluent prediction toward structured #thinking that can be inspected and trusted. Work coming out of #MBZUAI shows how open, sovereign models could shape the