Learning Functional Subspaces for Neural Network Compression Paper • 2609.40127 • Published 12 days ago • 25
TRIAGE: Direction-Aware Mismatch Stabilization of Native NVFP4 Reinforcement Learning Paper • 2610.07043 • Published 7 days ago • 42
World Editing: Intervening on Executable Worlds at Increasing Depth Paper • 2610.02331 • Published 11 days ago • 30
Groupwise Agentic Grading and Advantage Redistribution for Code Agent RL Paper • 2609.32577 • Published 16 days ago • 106
RewardVerse: Rubric-Guided Policy Optimization for Video Reward Modeling Paper • 2609.22947 • Published 23 days ago • 43
Think Before You Score: Thinking Reward Model for Visual Generation Paper • 2609.37372 • Published 13 days ago • 104
PanoVLN: Towards Effective Panoramic Vision-and-Language Navigation Paper • 2609.34759 • Published 14 days ago • 153
Surprising Success, Repeated Failure: Entropy-Guided Credit Assignment for Exploration in LLM Reasoning Paper • 2609.33781 • Published 15 days ago • 46
TRACE: Temporal Audit and Condition-aware Evaluation of Streaming Video Understanding Paper • 2609.30670 • Published 17 days ago • 11
IterSynth: Rethinking Deep Search Agents via Role-Decoupled Iterative Synthesis Paper • 2609.29444 • Published 18 days ago • 21
All modalities are equal, but video is more equal: Closing the Cross-Attention Gap in Joint Video Generation Paper • 2609.27901 • Published 19 days ago • 24
Rethinking Critic Learning in PPO: Understanding and Mitigating Value Flattening Paper • 2609.18708 • Published 26 days ago • 84
ShieldVLA: Feasibility-Aware Safety Alignment for Vision-Language-Action Models Paper • 2609.13231 • Published Sep 2 • 19
RULER: Instance-aware Rubric Rewards for SVG Generation Paper • 2609.25270 • Published 21 days ago • 104