Capability-Driven Self-Evolution of Agent Memory Paper • 2610.06361 • Published 6 days ago • 19
4DCodeBench: Benchmarking Agents on Inverse Graphics of Dynamic Scenes Paper • 2610.03715 • Published 9 days ago • 32
Adaptive Reward Routing: Dynamic Multi-Reward Optimization for Joint Audio-Video Diffusion via Forward-Process RL Paper • 2609.37200 • Published 12 days ago • 140
OSWorld-Science: A Benchmark of Computer Use Agents for Learning and Using Scientific Software Paper • 2609.39903 • Published 11 days ago • 64
HybridCUA: Learning to Orchestrate GUI and CLI for Computer-Use Agents Paper • 2609.38008 • Published 12 days ago • 48
Scaling Properties of Same-Family On-Policy Distillation Paper • 2609.32722 • Published 15 days ago • 326
MaLiang-Harness: A Programmable Path to Image and Video Generation Paper • 2609.34309 • Published 13 days ago • 394
Beyond Teacher Assignment: Domain-Normalized Multi-Teacher On-Policy Distillation Paper • 2609.35347 • Published 13 days ago • 180
TRACE: Temporal Audit and Condition-aware Evaluation of Streaming Video Understanding Paper • 2609.30670 • Published 16 days ago • 11