How much 3D do you actually see in a photo? 👀
A depth map stops at the visible surfaces. Our #ICML2026 paper LaRI (ruili3.github.io/lari) goes further: predicting the hidden geometry with layered point maps, in one feed-forward. (1/3)
Does 3D reconstruction have to be complex?
We answer this question with PointDiT (#ICML2026): a minimalist pixel-space Diffusion Transformer without bells and whistles.
We show that a plain ViT can estimate dense 3D point maps by operating directly on raw patches. No hybrid
🚀 The #ICCV2025 Award Candidate Papers are out! 🚀
From 2,701 submissions, only 13 were selected, spanning 3D vision, generative models, foundation models, and more.
Key highlights at a glance 👇
Another major update of the "awesome-dust3r" (github.com/ruili3/awesome…) paper list.
There are more VGG-T follow-ups and some interesting correlations, e.g., VGGT-Long/LONG3R, STream3R/Streaming 4D-VGGT.
Let's see what happens for visual geometry in the DINO-v3 era :)