AdaReasoner Dynamic Tool Orchestration for Iterative Visual Reasoning
@_akhaliq
-
High-Sparsity Attention Tuning for Video Diffusion Transformers
By
–
SALAD Achieve High-Sparsity Attention via Efficient Linear Attention Tuning for Diffusion Transformer
-
Jet-RL: On-Policy FP8 Reinforcement Learning with Unified Precision
By
–
Jet-RL Enabling On-Policy FP8 Reinforcement Learning with Unified Training and Rollout Precision Flow
-
SpaceTimePilot: Generative Rendering for Dynamic Scenes
By
–
SpaceTimePilot
— AK (@_akhaliq) 3 janvier 2026
Generative Rendering of Dynamic Scenes Across Space and Time pic.twitter.com/d4pHH1i53JSpaceTimePilot Generative Rendering of Dynamic Scenes Across Space and Time
-

FlowBlending: Stage-Aware Multi-Model Sampling for Video Generation
By
–
FlowBlending Stage-Aware Multi-Model Sampling for Fast and High-Fidelity Generation
-

Dream2Flow: Video Generation with 3D Object Flow for Manipulation
By
–
Dream2Flow
— AK (@_akhaliq) 2 janvier 2026
Bridging Video Generation and Open-World Manipulation with 3D Object Flow pic.twitter.com/5m09xRzjG0Dream2Flow Bridging Generation and Open-World Manipulation with 3D Object Flow
-
Hypergraph Memory Enhances Multi-step RAG Long-Context Modeling
By
–
Improving Multi-step RAG with Hypergraph-based Memory for Long-Context Complex Relational Modeling
-
DiffThinker: Generative Multimodal Reasoning with Diffusion Models
By
–
DiffThinker Towards Generative Multimodal Reasoning with Diffusion Models
-
Dynamic Large Concept Models with Latent Reasoning in Adaptive Semantic Space
By
–
Dynamic Large Concept Models Latent Reasoning in an Adaptive Semantic Space
-
JavisGPT: Multi-modal LLM for Video Understanding and Generation
By
–
JavisGPT A Unified Multi-modal LLM for Sounding-Video Comprehension and Generation