Prism: Dynamic Sparse Attention for Native 2K Joint Video-Audio Generation Model Training Paper • 2610.05416 • Published 8 days ago • 20
ANTMAN: Adaptive Need Tracking for Multi-Agent Navigation in Large Information Spaces Paper • 2609.33326 • Published 15 days ago • 37
All modalities are equal, but video is more equal: Closing the Cross-Attention Gap in Joint Video Generation Paper • 2609.27901 • Published 19 days ago • 24
HARMONY: Hierarchical Agentic Reasoning for MONocular Image-to-Scene Synthesis Paper • 2609.26793 • Published 20 days ago • 9
Harness-Zero: Harness Distillation via Agent-as-Harness Paper • 2609.24974 • Published 21 days ago • 38
onPanda: Efficient Annotation of On-Policy Alignment Data for LLMs and Agents via Token-Level Correction Paper • 2609.24983 • Published 21 days ago • 55
CARE: Experience-Guided Atomic Corrective Execution for Vision-Language-Action Policies Paper • 2609.24118 • Published 21 days ago • 28
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses Paper • 2609.24972 • Published 21 days ago • 225
DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression Paper • 2609.19969 • Published 25 days ago • 228
PANORAMA: Panoptic Grounded Captioning via Mask Proposal Selection Paper • 2609.19143 • Published 26 days ago • 19