Kimi k2 is a terrifyingly good open-source model. Its really good at code, agents, and real-world tasks. It’s open, fast, and beats most LLMs in key benchmarks. I built 3 working apps with it in a day:
LLMS
-

Memory Terminology Convergence in AI Systems
By
–
good point, i agree that memory is a confusing name for this. but it seems to be converging on that as common language. (eg openai, claude, langchain)
-

LLMs Struggle With Coherent Adventure Generation and Self-Learning
By
–
Historically, LLMs have trouble building a coherent adventure because there are many pieces that fit together. I also thought it was interesting that it researched how to write an adventure first, and followed the rules it discovered.
-

OpenAI’s Significant Statement About New Model
By
–
if youre not used to reading oai messaging you might miss how big a statement these guys are making about the new model
-

ChatGPT Agent Launch: New Frontier Model Unveiled Today
By
–
a lot of people are pooh poohing the ChatGPT Agent launch as “just a better harness” but they’re not reading closely enough we got a new frontier model today folks these charts are like for like, same harness they basically stopped short of calling it GPT5 but yeah if there
-
Foundation Models for NP-Hard Optimization Problems
By
–
See our recent work on applying foundation models to tackle challenging NP-Hard optimization problems:https://t.co/uX2YFnqsOYhttps://t.co/MDPwf8YTkx
— hardmaru (@hardmaru) 18 juillet 2025See our recent work on applying foundation models to tackle challenging NP-Hard optimization problems: https://
sakana.ai/ale-bench -
Model Misinterpretation: When AI Behavior Gets Misread as Core Truth
By
–
Yeah. seems it remembered what he told it, then when it said it back to him later he misinterpreted that as evidence of it being a core truth baked into the model.. :/
-
Relevant Memory in AI: Emerging Practices and Experimentation
By
–
oh yeah i’m letting the phrase “relevant memory” do a lot of work here lots of great experiments but don’t feel like there’s a “best practice” or “standard” emerging here yet
-

Upstage AI launches Solar Pro 2 reasoning model at competitive pricing
By
–
Super happy to see #solarpro2 is positioned among the top models in artificial analysis. Check it out at chat.upstage.ai. Artificial Analysis (@ArtificialAnlys) 🇰🇷 South Korean AI Lab Upstage AI has just launched their first reasoning model – Solar Pro 2! The 31B parameter model demonstrates impressive performance for its size, with intelligence approaching Claude 4 Sonnet in 'Thinking' mode and is priced very competitively Key details: ➤ Hybrid reasoning: The model offers optionality between 'reasoning' mode and standard non-reasoning mode ➤ Korean-language ability & Sovereign AI: Based in Korea, Upstage announced superior performance in Korean language evaluations. This release aligns with countries' interests to develop sovereign AI capabilities ➤ Pricing: Competitively priced at $0.5/1M tokens (input & output), significantly cheaper than comparable models including Claude 4 Sonnet Thinking ($3/$15/M input/output tokens) and Magistral Small ($0.5/$1.5/M input/output tokens) ➤ Proprietary: @upstageai has not released the model weights, though they have open-sourced previous Solar Pro models. Whether they will release Solar Pro 2's weights remains unclear as it wasn't mentioned in their announcement — https://nitter.net/ArtificialAnlys/status/1945961441888231709#m
-

South Korea’s AI Labs Challenge Global Leaders with EXAONE 4.0
By
–
South Korea is steadily climbing the global AI leaderboard. Despite having fewer GPUs and less infrastructure than China, Korean labs are beginning to make a strong impact. LG AI Research recently released EXAONE 4.0 (32B), which outperforms Qwen 235B in coding tasks and even