OneStreamer: Unifying Perception, Memory, and Proactive Response in Streaming Video Interaction Paper • 2610.01762 • Published 8 days ago • 234
Adaptive Reward Routing: Dynamic Multi-Reward Optimization for Joint Audio-Video Diffusion via Forward-Process RL Paper • 2609.37200 • Published 10 days ago • 138
A Missing Piece for Trustworthy AI Reviewers: From Benchmarking Rhetorical Robustness to SciCore Review Paper • 2609.39027 • Published 9 days ago • 76
ActiveSaddler: Automated Curriculum Learning for Agent Harness Optimization Paper • 2610.00906 • Published 8 days ago • 82
On-Policy or Off-Policy Learning? A Systematic Study of Distillation Dynamics Paper • 2609.35259 • Published 11 days ago • 198
Periodic Weak Spots: Phase Sensitivity from Chunked KV-Cache Compression Paper • 2609.36322 • Published 11 days ago • 112
What Makes World Action Models Generalize? An Empirical Study of Test-Time Future Modeling Paper • 2609.34981 • Published 10 days ago • 136
VoxMem: Benchmarking Multimodal Memory in Large Audio Language Models Paper • 2609.32607 • Published 13 days ago • 154
Scaling Properties of Same-Family On-Policy Distillation Paper • 2609.32722 • Published 13 days ago • 325
MaLiang-Harness: A Programmable Path to Image and Video Generation Paper • 2609.34309 • Published 11 days ago • 412
In-Context Learning for Robots: Methods and Applications Paper • 2609.36012 • Published 11 days ago • 390
Raven: The Harness of Harnesses for Composable Agentic Intelligence Paper • 2609.33439 • Published 12 days ago • 569
Duplex-MPE: Benchmarking Multi-Party Interaction in Full-Duplex Dialogue Paper • 2609.31948 • Published 14 days ago • 89
Self-Evolving Coding Agents: From Digital Programs to Physical-World Intelligence Paper • 2609.35432 • Published 11 days ago • 109
Groupwise Agentic Grading and Advantage Redistribution for Code Agent RL Paper • 2609.32577 • Published 13 days ago • 133