ViT-Up: Faithful Feature Upsampling for Vision Transformers Paper • 2606.14024 • Published Jun 12 • 10
AnyFlow: Any-Step Video Diffusion Model with On-Policy Flow Map Distillation Paper • 2605.13724 • Published May 13 • 104
SpatialClaw: Rethinking Action Interface for Agentic Spatial Reasoning Paper • 2606.13673 • Published Jun 11 • 115
VoLo: A Physical Orchestrator for Open-Vocabulary Long-Horizon Manipulation Paper • 2606.07723 • Published Jun 5 • 6
NVIDIA Nemotron v3 Collection Open, Production-ready Enterprise Models • 33 items • Updated Aug 14 • 378