RubricEM Meta-RL with Rubric-guided Policy Decomposition beyond Verifiable Rewards
@_akhaliq
-

EgoMemReason: A Memory-Driven Benchmark for Egocentric Video Understanding
By
–
EgoMemReason A Memory-Driven Reasoning Benchmark for Long-Horizon Egocentric Understanding
-

SenseNova-U1: Unifying Multimodal Understanding and Generation
By
–
SenseNova-U1 Unifying Multimodal Understanding and Generation with NEO-unify Architecture
-

Multi-Agent Synergy for Scaling Test-Time Compute
By
–
TMAS Scaling Test-Time Compute via Multi-Agent Synergy
-

Rebellious Student: Reversing Teacher Signals
By
–
Rebellious Student Reversing Teacher Signals for Reasoning Exploration with Self-Distilled RLVR
-

Multi-Agent Synergy for Scaling Test-Time Compute
By
–
TMAS Scaling Test-Time Compute via Multi-Agent Synergy
-

Soohak: A Benchmark for Evaluating Research-level Math in LLMs
By
–
Soohak A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs
-
Sharing Papers and Apps on Hugging Face
By
–
Paper:
https://huggingface.co/papers/2605.10922
…
App:
https://huggingface.co/spaces/TencentARC/Pixal3D
… -
Pixel-Aligned 3D Generation
By
–
Pixal3D
— AK (@_akhaliq) 12 mai 2026
Pixel-Aligned 3D Generation from Images pic.twitter.com/DADl4UQVIGPixal3D Pixel-Aligned 3D Generation from Images

