and yes it is a hallucination, despite what you say
LLMS
-
Version 5.3 Should Search Online Without Explicit Instructions
By
–
a friend sent it, 5.3, and you shouldn’t have to TELL it to search online.
-

Per-layer embeddings likely unused in final model implementation
By
–
I saw the per-layer embeddings in the code, but I don't think they were used in the final models. Maybe it was a left-over from some internal experiments.
-
LLM Reliability Issues in Mission Critical Applications
By
–
as i have said a million times, their unreliability is a serious issue in many mission critical applications. you can call that “probabilistic” but the problem remains. and it’s a bit of an abuse of the term. they aren’t calculating and reporting probabilities. and they didn’t
-
Neurosymbolic Systems Win Over Pure LLMs
By
–
neurosymbolic harnesses and tooling for the win. not pure LLMs, which is what my predictions were always about.
-

The AI Skill That Will Make You Billionaire
By
–
This skill will make you a billionaire. pic.twitter.com/5Y711wgNsh
— Anndy Lian (@anndylian) 3 avril 2026This skill will make you a billionaire.
-
τ³-bench: Interactive Agent Evaluations for Knowledge and Voice
By
–
Really excited for the release of 𝜏³-bench, which brings interactive agent evals ever closer to real-world use cases across two dimensions: 1. 𝜏-knowledge evaluates agents that need to operate over noisy knowledge bases to figure out the correct policies/tools to use while serving a user 2. 𝜏-voice tests voice agents in interactive customer service style settings. If you are developing embedding or voice models for AI agents, 𝜏³ is a great testbed for you to see how your models would perform in a realistic downstream use case. Blog: sierra.ai/blog/bench-advanci… Tweets from @BenShi34 and @keshav_57: nitter.net/benshi34/status/203436… nitter.net/keshav_57/status/20346…
-

Google DeepMind Adopts Apache 2.0 License for Open Models
By
–
Love that Google DeepMind is following OpenAI’s suit w/ using Apache 2.0 license for their open weights models – congrats! but, can we please stop using Arena Elo as the de facto measure of performance?
-

DeepSeek V4 Coming Soon with Native Huawei Ascend 950PR Support
By
–
Huge: DeepSeek v4 probably in the next few weeks – and it will be running natively on Huawei's Ascend 950PR Chips DeepSeek is about to drop its next-gen V4 model (via The Information) and for the first time, it'll run natively on Huawei's Ascend 950PR chips, marking a genuine
