An eventful day, but I am not personally worried for the following reasons: The new DeepSeek models show a number of advances that make training more efficient, and I expect those to be quickly incorporated by the leading US labs, giving them some additional performance
LLMS
-

DeepSeek Releases Janus-Pro-7B Open-Source Multimodal AI Model
By
–
DeepSeek just dropped ANOTHER open-source AI model, Janus-Pro-7B. It's multimodal (can generate images) and beats OpenAI's DALL-E 3 and Stable Diffusion across GenEval and DPG-Bench benchmarks. This is just super cool.
Now they will move their side project as their main project -
DeepSeek FAQ: Essential Reading on AI Development
By
–
Suggest reading (which is all I'm doing): https://
stratechery.com/2025/deepseek-
faq/
… -
Compute Power Critical for Next Generation AI Models
By
–
but mostly we are excited to continue to execute on our research roadmap and believe more compute is more important now than ever before to succeed at our mission. the world is going to want to use a LOT of ai, and really be quite amazed by the next gen models coming.
-
DeepSeek R1 Competition Drives Model Release Strategy
By
–
deepseek's r1 is an impressive model, particularly around what they're able to deliver for the price. we will obviously deliver much better models and also it's legit invigorating to have a new competitor! we will pull up some releases.
-

Building Custom Tools with OpenAI o1 for SWE-Bench
By
–
🆕 short pod – how @shawnup got SOTA SWE-Bench Verified
— Latent.Space (@latentspacepod) 28 janvier 2025
with @openai o1… by building his own tools (on @weights_biases Weave) to look at data! https://t.co/clyqvZNhHQ pic.twitter.com/ALmUSSAjiJshort pod – how @shawnup got SOTA SWE-Bench Verified with @openai o1… by building his own tools (on @wandb Weave) to look at data!
-

How DeepSeek R1 Reasoning Model Works
By
–
This is how DeepSeek r1 reasoning model works behind the scenes. A 3-turn chain of thought reflecting on 3 questions.
-

Nature publishes evolutionary model merging optimization research
By
–
論文「Evolutionary Optimization of Model Merging Recipes」が論文誌「Nature Machine Intelligence」に採択され本日掲載されました。最新バージョンでは本アプローチをさらに実証する新たな実験結果を含んでいます。ぜひ以下からご覧ください。 https://
nature.com/articles/s4225
6-024-00975-8
… Sakana -

Nature Publishes Evolutionary Optimization Model Merging Paper
By
–
We’re pleased to announce that our paper, “Evolutionary Optimization of Model Merging Recipes,” has been accepted to Nature Machine Intelligence and published today! The latest version includes new experimental results that further validate our approach. https://
nature.com/articles/s4225
6-024-00975-8
… -
Why Learning LLMs Is a Critical Skill for Everyone
By
–
I know LLMs can feel intimidating at first, but the rewards are worth it. Whether you’re a developer, a business leader, or just someone curious about the future, learning how to use these tools is a skill that’ll pay off in ways you can’t even imagine yet. So… what’s stopping