metaai·lightalo unofficial · independent
Universe / Vision / VGGT
Vision · 2025

VGGT

VGGT (Visual Geometry Grounded Transformer)

CVPR 2025 Best Paper: geometry without the geometry pipeline

open source 1B params
★ 14k vggt❝ 1.7k citationsread 2026-09-03

Latest: VGGT (CVPR 2025)

CVPR 2025 Best Paper winner from Meta AI and Oxford's Visual Geometry Group: a single feed-forward transformer that infers camera parameters, depth maps, point maps, and 3D point tracks from one to hundreds of images in seconds — replacing whole classical structure-from-motion pipelines with one network pass. Code is open source.

Why it matters

VGGT replaced the classical structure-from-motion pipeline with one feed-forward transformer that infers cameras, depth, point maps and tracks from one to hundreds of images in seconds. Chosen Best Paper from over 13,000 CVPR 2025 submissions, it reset expectations for how much 3D geometry a single network pass can recover.

Facts

Try it yourself

Sources

More in Vision

SAM 3SAM 3D (Objects + Body)DINOv3Perception Encoder & Perception Language ModelSAM 2Chameleon

Read the Vision story on the sky →

✦ Open on the map Explore Vision Quiz me