DiffThinker Towards Generative Multimodal Reasoning with Diffusion Models
MULTIMODAL AI
-
JavisGPT: Multi-modal LLM for Video Understanding and Generation
By
–
JavisGPT A Unified Multi-modal LLM for Sounding-Video Comprehension and Generation
-
AI competition in synthetic pornography creation
By
–
Dans les 6 prochaines années nous allons voir émerger une compétition extrêmement serrée entre des IA dégénératives qui produisent du contenu pornographique à partir de photos et vidéos réelles de vrais humains.
-
DeepSeek’s 2026 hyper-connection breakthrough
By
–
The whale is back DeepSeek dropping manifold-projected hyper-connections right after the holidays is the kind of energy we needed for 2026. Stabilizing those long-range skips by enforcing geometric alignment instead of letting them warp the representation space elegant fix
-

IPA: Chunk-Level Credit Assignment Breakthrough
By
–
Here's the breakthrough nobody expected: They introduced IPA (Interaction-Perceptive Agentic Policy Optimization) that assigns credit at the CHUNK level, not token level. Tokens are too fine. Trajectories are too coarse. Chunks align with actual tool-use semantics.
-

AI Infrastructure Breakthrough in 2025
By
–
Chinese AI labs just dropped a bombshell research paper that exposes why 99% of "AI agent" companies are building on broken infrastructure. The ROME model + ALE ecosystem might be the most important open-source release of 2025. Here's what nobody's talking about:
-
Kling v2.6 Released: Advanced Text-to-Video and Image-to-Video Generation
By
–
Kling v2.6 is here
— Replicate (@replicate) 31 décembre 2025
Generate cinematic text-to-video and image-to-video clips with fluid motion, photorealistic details, and native audio pic.twitter.com/tZRJgCBTn5Kling v2.6 is here Generate cinematic text-to-video and image-to-video clips with fluid motion, photorealistic details, and native audio
-

Qwen Image 2512 Launched with Improved Realism and Speed
By
–
Qwen-Image-2512 is here Qwen Image 2512 is an improved version of Qwen Image with more realistic human generation, finer textures, and stronger text rendering We've worked with @PrunaAI again to deliver you the fastest speeds possible
-
DreamOmni3: Scribble-Based Editing and Generation Tool
By
–
DreamOmni3 Scribble-based Editing and Generation
-

UltraShape 1.0: High-Fidelity 3D Shape Generation via Geometric Refinement
By
–
UltraShape 1.0 High-Fidelity 3D Shape Generation via Scalable Geometric Refinement