“Scalable Training of Mixture-of-Experts Models with Megatron Core” This NVIDIA MoE report walks through the hard part of MoE training. The key is not to add more parameters, but keeping sparse models efficient when only a small part of the model runs for each token. For
COMPUTING
-

CUDA Agent: Large-Scale Agentic RL for Kernel Generation
By
–
CUDA Agent | Large-Scale Agentic RL for CUDA Kernel Generation https://
buff.ly/13zVaiQ
#AI #MachineLearning #DeepLearning #LLMs #DataScience -
Token Generation Speed Comparison: OSS and Nemotron Models Performance
By
–
Yeah. Not quite as bad for me, but:
gpt-oss-120b: 42.22 tokens/sec
nemotron-3-super: 20.43 tokens/sec on the DGX Spark. But this is ollama and might be an implementation issue. I have yet to try the Nvidia-optimized llama.cpp version (they had one for Nemotron Nano back then) -
One-Shot iOS App Development: From Months to Minutes
By
–
I remember coding my first very basic iOS app in months haha. The fact that this is one-shot now is insane.
-

Steve Jobs’ Vision and Early Apple Struggles with Personal Computers
By
–
Last month, @StanfordHAI leaders @SuryaGanguli and @drfeifei shaped critical conversations through keynotes and panels at the India AI Impact Summit 2026. This global presence further underscores our mission in bridging AI innovation and governance across global contexts.
-

llmfit: Auto-detect hardware and rank 206 models by VRAM compatibility
By
–
Stop guessing which models fit in your VRAM! llmfit is a CLI tool that auto-detects your hardware and ranks 206 models by what actually runs on your system. You download a 70B model and hope it fits. Or you estimate memory requirements across quantization levels and still end
-
India launches AI sovereignty with 38,000 GPUs infrastructure
By
–
This is what sovereignty looks like. Not just a word in a speech. 38,000 GPUs. 12 foundational model teams. 10000+ datasets. India's AI era has started. And YOU are in it. #MadeWithIndiaAI (8/8) @AshwiniVaishnaw @jitinprasada @PIB_India @SecretaryMEITY @abhish18 @kavitabha
-
India Leads Global AI Skill Penetration Rankings with Compute Access
By
–
India ranks #1 globally in AI skill penetration (Stanford AI Index 2024). Score of 2.8, ahead of the US at 2.2. We have the talent. Now we have the compute. The only direction is forward. (7/8) @AshwiniVaishnaw @jitinprasada @PIB_India @SecretaryMEITY @abhish18 @kavitabha
-
India Launches National Compute Pool with Seven Cloud Providers
By
–
In 18 months, seven empanelled cloud providers. @NVIDIA H100s, @AMD MI300s, @Google Trillium TPUs. Distributed across tier-3 data centres. All under one national compute pool. (3/8) @AshwiniVaishnaw @jitinprasada @PIB_India @SecretaryMEITY @abhish18 @kavitabha @GoI_MeitY
-
India Launches IndiaAI Mission with ₹10,372 Crore Sovereign Infrastructure
By
–
March 2024>> The IndiaAI Mission launches with ₹10,372 crore. Goal: Give India sovereign AI infrastructure. Not borrow it – or not rent it from abroad… BUILD it! (2/8) @AshwiniVaishnaw @jitinprasada @PIB_India @SecretaryMEITY @abhish18 @kavitabha @GoI_MeitY @_DigitalIndia
