AI Dynamics

Global AI News Aggregator

About

TOOLS

  • Building Weather Apps with X API and Replit
    Building Weather Apps with X API and Replit

    Met a founding engineer today from @Replit
    . Jen Li. We were both judging the @Pokee_AI hackathon. They have me some credits and I built two apps in 20 minutes using the X API:a weather one which mapped storms being reported in by my climate scientist list and another monitoring

    → View original post on X — @scobleizer

  • Karpathy’s LLM Knowledge Base System and the Future of Memory Infrastructure

    Karpathy posted a long thread about his most frequent use cases with LLMs recently. Not writing code, but building knowledge bases. The approach is quite hardcore: he dumps papers, articles, code repositories and other materials into a folder, then lets an LLM "compile" them into a Markdown wiki. The wiki includes summaries, backlinks, concept categorization, and articles linked to each other. The frontend uses Obsidian for viewing, and Q&A also has the LLM retrieve against the wiki. In his own words, most token consumption now isn't in manipulating code, but in manipulating knowledge. This shift is quite interesting. The entire system can also maintain itself. He wrote some LLM "health check" scripts that periodically scan the wiki for contradictory data, missing information, and potential connections, letting the LLM patch itself. The results from each Q&A can also be archived back into the wiki, making it thicker with each use. Actually, Karpathy clarified something that's happening right now: the greatest value of LLMs might not be helping you generate content, but helping you manage knowledge. But the last sentence of his post is the most worth pondering: "I think there is room here for an incredible new product instead of a hacky collection of scripts." He himself knows this system is hacked together from scripts. Obsidian + command line + manual processes—it works, but it's just a demo. And there are several problems he probably felt:
    The wiki is local Markdown files, tied to the computer—it breaks when you switch machines. Retrieval relies on the LLM's own maintained indexes and summaries; he said around 400K words it still holds up, but beyond that? He even said "I thought I had to reach for fancy RAG," just that the scale hasn't reached that point yet. A more fundamental problem is that the wiki stores knowledge, but not memory. What does that mean? Knowledge is "domain X has these concepts, and their relationships are like this." Memory is "I just read a paper last week that refutes this viewpoint, and my judgment on this direction changed." One is static, one walks with you. Karpathy's system can help you store things and search things, but it doesn't know you've changed.
    This is actually the difference between a knowledge base and a memory system. The gap isn't a better script—it's an entire architecture. The model can't just "store" and "search"; it needs to sense which information is relevant to who you are now, needs to evolve itself as you use it, needs to maintain coherence across projects and timelines. Karpathy proved with a hand-rolled solution that this direction is right. But he also proved firsthand that you can't go far with just file systems and prompts. Memory needs to be infrastructure, not a collection of scripts. [Translated from EN to English]

    → View original post on X — @elliotchen100, 2026-04-05 04:31 UTC

  • README-Driven Development with Claude Code for Tool Building

    I built this one using README-driven-development: I hand crafted a detailed README describing exactly how the tool should work… then dumped that into Claude Code and told it to build it gisthost.github.io/?d4b1a398…

    → View original post on X — @simonw, 2026-04-05 04:20 UTC

  • Codex App Server Enables Easy Agentic App Development

    Codex app server makes it easy to build your own agentic apps: am.will (@LLMJunky) The Codex app server was such a brilliant stroke of foresight that really doesn't get enough love Not only are you allowed to use your chatgpt account with any harness, but you can build your own apps directly on top of theirs. They just make building on and with codex such a great experience To demonstrate this utility, I want to highlight the kitty litter app, made by @SIGKITTEN. Instead of having to build the entire harness, and all the infrastructure, he's plugged into the app server for a unified experience between mobile and dev machine. When I create a session on my computer, it's automatically available on my phone. All of the chats you see in this video automatically populated when we connected to the app server. All my skills. My agents. My sessions. My folders. My prompts. They're all ready to use – automatically. Because they're exposed by the app server, along with many other endpoints. It's a great ux/dx that really deserves some love. It's almost like they want you to build on top of their products 😉 Btw Litter is great 👍 — https://nitter.net/LLMJunky/status/2040506388292546761#m

    → View original post on X — @gdb, 2026-04-05 03:18 UTC

  • Codex and Claude Agent SDK Server Limitations Discussed

    Codex app server is limited in its own ways, unfortunately – it does too much server-side which limits what one can do with it. (Claude Agent SDK has it's own silly limitations too; no-one has actually got this right yet.)

    → View original post on X — @jeremyphoward

  • Gemma 4-31B Ranks #27, Outperforming GPT-5.2 and Qwen3.5
    Gemma 4-31B Ranks #27, Outperforming GPT-5.2 and Qwen3.5

    After thousands of additional votes in head-to-head competition, Gemma4-31B continues to score #27 ahead of Qwen3.5-397b, Ernie-5, and GPT-5.2-high. David Hendrickson (@TeksEdge) 🚨 Big 📏 News! @Google's new Gemma 4 -31B (dense) and Gemma 4-27B A4B (MoE) scored #27 and #40, respectively on @arena. 🦾 Dense #⃣ > than MoE! 👀 👇 🎯 Gemma 4-31B scores better than; ☑️GPT-5.1 ☑️GPT-5.2 ☑️GPT-5.2-high ☑️ Qwen3.5-2-397b-a17B 💳 Q4-RTX-5090 @ 50-80 tps? @JeffDean @demishassabis — https://nitter.net/TeksEdge/status/2040226061355774348#m

    → View original post on X — @deeplearn007, 2026-04-05 02:48 UTC

  • Microsoft’s Copilot Proliferation: 78 Products and Counting
    Microsoft’s Copilot Proliferation: 78 Products and Counting

    Microsoft has become Megabloat. Bearly AI (@bearlyai) Tey Bannerman counted up all the products and tools that Microsoft has named “Copilot”. Found 78 of them: “there are now Copilots inside Copilots, Copilots for other Copilots, and a physical Copilot key on your keyboard for summoning them.” — https://nitter.net/bearlyai/status/2040548562053050749#m

    → View original post on X — @pmddomingos, 2026-04-05 02:39 UTC

  • AI Tools Transform Family Healthcare Decision-Making During Crisis

    This isn’t an edge case. From anonymized U.S. ChatGPT data, we are seeing: • ~2M weekly messages on health insurance • ~600K weekly messages from people living in “hospital deserts” (30 min drive to nearest hospital) • 7 out of 10 msgs happen outside clinic hours Simon Smith (@_simonsmith) I’ve been critical of OpenAI lately, but for the past three weeks my family has been dealing with a health issue with my dad, and a ChatGPT shared project with live document syncing has been essential to organizing and understanding everything happening. Me, my four siblings, my mom, and my dad have faced an onslaught of information from various doctors and nurses, which we’ve captured in hundreds of text messages and documents and scans and you name it. ChatGPT has helped us collect this information in a single place, make sense of it, and interrogate it to make the most informed decisions possible. Also, credit where due: Claude played an important role as well, by ingesting iMessages and synthesizing summarizes from them to upload to ChatGPT, as well as by extracting text from a bunch of HEIC document scans. I think those of us, like me, excited at AI’s potential get frustrated when we can see issues so clearly, like ChatGPT’s bad design skills, and Claude’s increasing instability and confusing usage consumption. But at times like this I’m reminded of how incredible this technology already is, letting me and my family make sense and act on hundreds of pieces of information, empowering us in the face of a disjointed and fragmented healthcare system. — https://nitter.net/_simonsmith/status/2040539824034115676#m

    → View original post on X — @jeremyphoward, 2026-04-05 01:43 UTC

  • Liberate OpenClaw with Open and Local Models from Hugging Face

    Liberate your @openclaw with an open model or local model with these tools from our friend @ClementDelangue and team at @huggingface 🦞 huggingface.co/blog/liberate…

    → View original post on X — @ceobillionaire, 2026-04-05 01:08 UTC

  • Token Counting API Support Across Bedrock Vertex Azure Anthropic

    yep we do this for Bedrock, Vertex, and Azure, since they don’t have a token-counting API available yet. When using Anthropic API we use the token-counting endpoint directly

    → View original post on X — @bcherny