I would have thought that Grok is the most based model, i.e. most helpful in letting you think about edgy science stuff, most independent when discussing politics and policy etc., but it's not. Grok is a normie with an anime face
LLMS
-

Meta’s Continuous Chain-of-Thought Reasoning for LLMs
By
–
Can reasoning LLMs think better if their Chain-of-Thought is continuous instead of discrete? This Meta paper introduces the first scalable way to train continuous CoTs with reinforcement learning—no need to distill from discrete references. By using "soft" tokens
-
100+ Free Open Source AI Agents and RAG Tutorials
By
–
Stay tuned for more such interesting posts → @Saboo_Shubham_ I have created 100+ AI Agents and RAG tutorials, 100% free and opensource. P.S: Don't forget to star the repo to show your support
-
100+ Free Step-by-Step AI Agents and RAG Systems Tutorials
By
–
100+ free step-by-step tutorials with code covering: AI Agents RAG Systems Voice AI Agents MCP AI Agents Multi-agent Teams Autonomous Game Playing Agents P.S: Don't forget to subscribe for FREE to access future tutorials.
-

DeepSearch: Training Small Reasoning Models More Effectively
By
–
How do you train small reasoning models more effectively? Many AI developers run into the same problem: RL fine-tuning plateaus quickly, especially for 1–2B parameter models. A new approach called DeepSearch offers a neat solution. Instead of only using Monte Carlo Tree
-

Claude Sonnet 4.5: Memory Breakthrough for AI Coding
By
–
Claude Sonnet 4.5: The Real Breakthrough Is Memory This week, @AnthropicAI unveiled Claude Sonnet 4.5. The coding improvements are real — but the deeper shift isn’t about code at all. It’s about memory. Claude Code now includes a memory tool that can persist files to disk
-
Claude Sonnet 4.5 Advances Defensive Cybersecurity Skills
By
–
We’ve focused on improving Claude’s skills in defensive cybersecurity. The results of this are visible in Claude Sonnet 4.5, which is comparable or superior to Opus 4.1 in cybersecurity tasks—yet both faster and cheaper. Read more:
-
Google releases NanoBanana model for production via Gemini API
By
–
Here’s a quick look at what we shipped this week — @NanoBanana is now generally available and ready for production use for developers via the Gemini API on @GoogleAIStudio and @GoogleCloud Vertex AI for enterprise customers. — We also released 10 new aspect ratios and
-

Apple’s Learning to Reason Method for Detecting Hallucination Spans
By
–
Apple presents Learning to Reason for Hallucination Span Detection
-
Claude vs GPT-5: Performance Degradation in Extended Conversations
By
–
Claude struggled a bit with syntax errors after a long chat, uncharacteristic, but complexity did build up. Whereas GPT-5 just seems to stop working after a while though, it thinks and then says what it'd do without doing it.