GPT is panicking and making mistakes "Yes, that was wrong again: I panicked and recycled an already used news image. I'm now replacing both"
LLMS
-
User notices Codex quality decline, asks others
By
–
Oh, and btw, Codex quality has gotten noticeably worse. Is it just me, or have you been seeing the same decline in quality?
-

Types of LLMs in AI Agents by @ingliguori
By
–
Types of #LLMs in #AIAgents by @ingliguori #GenerativeAI #ArtificialIntelligence #MachineLearning #ML
-

EVE-Agent forces verifiable source spans for AI learning
By
–
Can your AI justify what it learns, or is it just guessing? Researchers from Fujitsu and the University of Tokyo present EVE-Agent: a self-evolving system that forces every training example to include a verifiable source span — no more learning from unsupported answers.
-
DeepSeek AI Economics: Margin, Pricing, and Cost Structure
By
–
Margin is the third lever. DeepSeek trained with no outside investment money and still sells API quota at a profit. They don't have demanding a return, which lets them price closer to actual cost.
-

AI Agents Beyond LLMs: Orchestration and Architecture
By
–
AI agents are more than LLMs • Skills → expertise
• MCP → external connections
• Subagents → delegated tasks
• Hooks → automation
• Memory → persistent context Real AI systems are orchestration layers. Via Giuliano Liguori (
@ingliguori
) #AI #AIAgents #MCP -

Abacus AI ChatLLM: Unified Interface for Frontier Models
By
–
Claude Opus 4.7, Sonnet 4.6, GPT-5.5 (Thinking + Pro), Gemini 3.1 Pro, DeepSeek V4 Pro, and Kimi 2.6 Thinking. All running inside Abacus AI ChatLLM – one workspace to access, compare, and switch between the top frontier models without juggling tools. Opus 4.7 → deep reasoning,
-
Gary Marcus says Claude Code works in neurosymbolic system but fails alone
By
–
it’s fine as a component in a larger (neurosymbolic) system, which is at last what Anthropic figured out, with Claude Code. but did in fact get stuck when used on its own, as i had predicted.
-

LLM agents with evolving episodic memory: CASCADE framework
By
–
What if your LLM kept learning after deployment—even without changing its weights? Researchers from Jilin University, King's College London, and UCL introduce CASCADE, a framework that equips LLM agents with an evolving episodic memory. It treats experience reuse like a
-

Codex Desktop missing context/token usage indicator?
By
–
Codex Desktop no longer shows visible context/token usage indicator? Bug or did they delete it?
