3/ Models in the lineup: Granite-4.0-H Small → 32B params (9B active), ideal for agents & support automation
Granite-4.0-H Tiny → 7B (1B active), low-latency + edge use
Granite-4.0-H Micro → 3B dense hybrid
Plus a 3B transformer-only Micro
COMPUTING
-
Granite-4.0-H Model Lineup: 32B Small to 3B Micro
By
–
-

Granite 4.0 Hybrid Mamba Transformer Cuts GPU Memory 70%
By
–
Granite 4.0 introduces a hybrid Mamba + transformer architecture. Cuts GPU memory needs by up to 70% Runs on cheaper hardware Faster inference, even with long contexts or multiple sessions
-
IBM Granite 4.0: Hybrid LLM for Enterprise Efficiency
By
–
🚨 IBM just dropped Granite 4.0 — hybrid Mamba/transformer LLMs built for enterprise efficiency.
— Futurepedia – Learn to Leverage AI (@futurepedia_io) 3 octobre 2025
✅ Up to 70% less RAM
✅ ISO 42001 certified & cryptographically signed
✅ Open-sourced + cheaper GPUs
Here’s why it matters 👇🧵 pic.twitter.com/4jtcWHzOwRIBM just dropped Granite 4.0 — hybrid Mamba/transformer LLMs built for enterprise efficiency. Up to 70% less RAM ISO 42001 certified & cryptographically signed Open-sourced + cheaper GPUs Here’s why it matters
-
GPU RTX 3090 Build and AI Streaming Setup Launch
By
–
we haven't even started > waiting on recording gear & 2 white desks > livestreams will be from office & basement > lighting is everything (very difficult) > 14x RTX 3090 build video will happen > Buy a GPU site next week > more on Local AI Summit soon > hopefully giveaways
-
Massive ping activity detected across 8 million pages
By
–
elle a observé des pings sur 8 millions de pages
-

Polya’s Timeless Approach Inspires Modern AI Research Methods
By
–
The approach we're building on is partly inspired by Polya's work from 80 years ago. The ideas are timeless!
-
Platform Consolidation: Enterprise AI Operations Across All Functions
By
–
We've refined the process and the platform. We now basically live in the platform—we do all our sysadmin work in it, host production apps in it, develop most of our software in it, our legal team does contract drafting in it, we iterate on GUIs in it, etc.
-

Turning Chat Discussions Into Scalable Deployment Plans
By
–
E.g before we built the scalable server farm for expanding the solveit platform, I chatted to my team about approaches we could take. You can see here a snippet of the dialog I used to turn that chat into a deployment plan—the transcript and style guide coming from variables.