4DAnyone: Create Anyone in 4D from a Casual Monocular Video Paper • 2608.20335 • Published Aug 20 • 87
GAE: Learning a Geometry-Native Latent Space for 3D-Consistent World Generation Paper • 2609.24981 • Published 15 days ago • 74
All modalities are equal, but video is more equal: Closing the Cross-Attention Gap in Joint Video Generation Paper • 2609.27901 • Published 13 days ago • 24
Rate-distortion optimization for full-reference image quality metrics via stochastic Hessian estimates Paper • 2609.30077 • Published 12 days ago • 7
TT-VidT: Decoupling the Temporal Axis for Efficient Motion-Centric Video Pretraining Paper • 2609.33419 • Published 9 days ago • 19
TrackEverything: Long Horizon Dense Tracking via De-Duplicating 3D Scene Representations Paper • 2609.30222 • Published 12 days ago • 14
VGGT-Diff: Visual Geometry Meets Diffusion for Sparse-View Novel View Synthesis Paper • 2609.33253 • Published 9 days ago • 5
EvolvingAvatar: Interactive 3D Head Generation That Adapts as Conversations Unfold Paper • 2609.35616 • Published 8 days ago • 4
GeoVerse: World-Consistent Novel View Synthesis in Geometric Latent Space Paper • 2609.35734 • Published 8 days ago • 9
In-Flight KV Cache with Clean Anchors for Faster Autoregressive Video Diffusion Paper • 2609.32540 • Published 10 days ago • 35
SpatialSpeak: QA-Native Reconstruction with Local and Global Context for Spatial Chain-of-Thought Reasoning Paper • 2609.33616 • Published 9 days ago • 27
LEGO-Anything: Coding Agents for 3D Scene Reconstruction Paper • 2609.36380 • Published 8 days ago • 138
TaoAvatar: Real-Time Lifelike Full-Body Talking Avatars for Augmented Reality via 3D Gaussian Splatting Paper • 2503.17032 • Published Mar 21, 2025 • 28
FFaceNeRF: Few-shot Face Editing in Neural Radiance Fields Paper • 2503.17095 • Published Mar 21, 2025 • 6
FRESA:Feedforward Reconstruction of Personalized Skinned Avatars from Few Images Paper • 2503.19207 • Published Mar 24, 2025 • 5
DiffPortrait360: Consistent Portrait Diffusion for 360 View Synthesis Paper • 2503.15667 • Published Mar 19, 2025 • 9
ChatAnyone: Stylized Real-time Portrait Video Generation with Hierarchical Motion Diffusion Model Paper • 2503.21144 • Published Mar 27, 2025 • 28