Droplet3D Commonsense Priors from Videos Facilitate 3D Generation
@_akhaliq
-

TiKMiX: Dynamic Data Influence Mixture for Language Model Pre-training
By
–
TiKMiX Take Data Influence into Dynamic Mixture for Language Model Pre-training
-

R-4B: Auto-Thinking MLLMs via Bi-Mode Annealing and RL
By
–
R-4B Incentivizing General-Purpose Auto-Thinking Capability in MLLMs via Bi-Mode Annealing and Reinforce Learning
-

EmbodiedOneVision: Vision-Text-Action Pretraining for Robot Control
By
–
EmbodiedOneVision Interleaved Vision-Text-Action Pretraining for General Robot Control
-
Building Video Transcription AI with Apple FastVLM Locally
By
–
vibe coding a video transcription AI app with Apple FastVLM on Hugging Face, a couple prompts in anycoder
— AK (@_akhaliq) 30 août 2025
100% locally in your browser (zero install) pic.twitter.com/cBhVrUfls6vibe coding a video transcription AI app with Apple FastVLM on Hugging Face, a couple prompts in anycoder 100% locally in your browser (zero install)
-

Microsoft rStar2-Agent Achieves State-of-the-Art Reasoning Performance
By
–
Microsoft presents rStar2-Agent Agentic Reasoning Technical Report rStar2-Agent boosts a pre-trained 14B model to state of the art in only 510 RL steps within one week, achieving average pass@1 scores of 80.6% on AIME24 and 69.8% on AIME25, surpassing DeepSeek-R1 (671B) with
-
Mixture of Contexts for Long Video Generation
By
–
Mixture of Contexts for Long Video Generation pic.twitter.com/vRVRhos8Ei
— AK (@_akhaliq) 29 août 2025Mixture of Contexts for Long Generation
-

Building NVIDIA Nemotron Nano Chat App with AnyCoder
By
–
vibe coding a NVIDIA-Nemotron-Nano-9B-v2 chat app in anycoder only took a couple a prompts and deployed with zero-gpu on Hugging Face NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning
-
MiniCPM-V 4.5 Chat App Development with Impressive Benchmark Performance
By
–
vibe coding a MiniCPM-V 4.5 @OpenBMB chat app in anycoder
— AK (@_akhaliq) 28 août 2025
MiniCPM-V 4.5 achieves an average score of 77.0 on OpenCompass, a comprehensive evaluation of 8 popular benchmarks. With only 8B parameters, it surpasses widely used proprietary models like GPT-4o-latest, Gemini-2.0 Pro,… pic.twitter.com/r1i5b6JfFpvibe coding a MiniCPM-V 4.5 @OpenBMB chat app in anycoder MiniCPM-V 4.5 achieves an average score of 77.0 on OpenCompass, a comprehensive evaluation of 8 popular benchmarks. With only 8B parameters, it surpasses widely used proprietary models like GPT-4o-latest, Gemini-2.0 Pro,

