LMCache an LLM serving engine extension to reduce TTFT and increase throughput, especially under long-context scenarios
LLMS
-
Skepticism about AI-writing detection systems
By
–
(I should add I haven’t actually tried Pangram until now; I’m just skeptical a priori of any system that claims to detect AI writing. Based on this, maybe Pangram is the exception.)
-
XML Prompts Remain Superior for Serious LLM Practitioners
By
–
Sorry, but it's not. I've been prompting/fine-tuning LLMs since 2019, and have tried just about everything, ran tests, etc. etc. There's a reason most serious practitioners use XML prompts for almost everything.
-
GDB Talk and Latent Space Podcast Questions for Greg
By
–
also hit the bell to summon algo gods for the gdb talk pls thx! https://
youtube.com/watch?v=avWhre
BUYF0
… we're taking suggestions and questions for the @latentspacepod with him, if you have Open Model or GPT5 or anything AI Engineering + OpenAI that you specifically want to hear from greg, -
Tool Call ID Mismatch Errors Fixed in AI Framework
By
–
It's SO good, and doesn't break with annoying tool call ID mismatch errors
-
GPT-5 reportedly being tested in ChatGPT platform
By
–
Possibly GPT-5 being tested in ChatGPT https://t.co/gb4UUQXjXh
— Peter Gostev (@petergostev) 28 juillet 2025Possibly GPT-5 being tested in ChatGPT
-
Optimized Prompts for Advanced AI Models
By
–
Oh, I meant that that looked like an optimized-for-o3 prompt!
-
Claude Models Performance Evaluation by Serious Practitioners
By
–
almost every model, especially claude models try it. almost any serious practitioner will tell you the same.
-
Structured Tags in AI: From Simple to Complex Use Cases
By
–
Our use cases range from extremely simple (i.e. output three options in tags) to absurdly complex (super nested tags with counters, metadata, etc.) Prompt clearly, and it'll work.
-

Kimi K2 and Qwen 3 Coder Impact on LLM Market Share
By
–
Impact of Kimi K2 and Qwen 3 Coder on the LLM market, based on the @openrouter data in the 'programming' category. What we see is quite interesting: – Sonnet 4 models keep growing as if nothing happened – Gemini 2.5 Pro is losing share very quickly, from 15% to 9% in a
