6). Star Attention: Efficient LLM Inference over Long Sequences – introduces Star Attention, a two-phase attention mechanism that processes long sequences by combining blockwise-local attention for context encoding with sequence-global attention for query processing and token
LLMS
-

o1 Replication: Distillation and Fine-tuning for Math Reasoning
By
–
3). o1 Replication Journey – Part 2 – shows that combining simple distillation from o1's API with supervised fine-tuning significantly boosts performance on complex math reasoning tasks…
-

LLM-Brained GUI Agents: Survey of Techniques and Applications
By
–
4). LLM-Brained GUI Agents – presents a survey of LLM-brained GUI Agents, including techniques and applications.
-
LLM Agents with Anthropic: Actionable Insights from Latent Space
By
–
This was a great episode featuring @ErikSchluntz with lots of actionable take-aways. I wrote what I found most interesting here, but it’s worth a full listen. @latentspacepod is excellent and you’re missing out if you don’t check it out : https://
mattstockton.com/2024/11/29/llm
-agents-anthropic.html
… -

Diffusion Models: Key Perspectives in Modern AI
By
–
Perspectives on diffusion https://
bit.ly/45nzZjn
#AI #MachineLearning #DeepLearning #LLMs #DataScience -

AI Giants Battle: Mistral, Amazon, Anthropic Shape 2026 Landscape
By
–
𝐋𝐞 𝐓𝐬𝐚𝐫 𝐆𝐏𝐓 𝐯𝐚-𝐭-𝐢𝐥 𝐩𝐞𝐫𝐝𝐫𝐞 𝐬𝐚 𝐜𝐨𝐮𝐫𝐨𝐧𝐧𝐞? Mistral débarque à Palo Alto Amazon dégaine Olympus Anthropic lance le Model Context Protocol Trump veut nommer un Tsar de l’IA Et bien plus encore sont au menu de la 18e édition de 𝕄𝕖𝕤
-

Complete Data Science Learning Resource Handbook Guide
By
–
GitHub – andresvourakis/data-scientist-handbook: This is a repo with links to everything you'd ever want to learn about data science https://
bit.ly/3N6nZvo
#AI #MachineLearning #DeepLearning #LLMs #DataScience -

Embedding Models in RAG Pipelines: Essential Guide
By
–
Good morning everyone! Today, we dive into an important part of the Retrieval-Augmented Generation (RAG) pipeline: the embedding model. All the data you have will be entered into embeddings, which we’ll then use to retrieve information. So, it’s quite important to understand
-

Model falsely claiming to be Amazon Titan
By
–
There also seems to be a model which is claiming to be Amazon Titan
-

New Google models on LMSYS, Goblin = Ultra?
By
–
A bunch of new models from Google have been added to @lmarena_ai Goblin = Ultra? Is it you?