[CVPR 2025 Best Paper Award] VGGT
Photorealistic Synthetic Dataset for Holistic Indoor Scene
Codex plugin that turns attached object images into code-only
Tiny vision language model
Models for object and human mesh reconstruction
A Unified Framework for Text-to-3D and Image-to-3D Generation
Unifying 3D Mesh Generation with Language Models
From Images to High-Fidelity 3D Assets
Rebuild the object in a reference image as a code-only, procedural
An extensive node suite that enables ComfyUI to process 3D inputs
Sharp Monocular Metric Depth in Less Than a Second
Code to accompany "A Method for Animating Children's Drawings"
MapAnything: Universal Feed-Forward Metric 3D Reconstruction
Navigation mesh generation and pathfinding toolkit for game AI systems
High-Resolution 3D Human Digitization from A Single Image
Code for "Image Generation from Scene Graphs", Johnson et al, CVPR 201
Computer vision and image processing library for Qt.