Python inference and LoRA trainer package for the LTX-2 audio–video
Native and Compact Structured Latents for 3D Generation
Fast stable diffusion on CPU and AI PC
code for Mesh R-CNN, ICCV 2019
VGGSfM: Visual Geometry Grounded Deep Structure From Motion
Generating Immersive, Explorable, and Interactive 3D Worlds
Accurate × Fast × Comprehensive
Tiny vision language model
Open image model at the forefront of design
Fast and Universal 3D reconstruction model for versatile tasks
AI Suite for upscaling, interpolating & restoring images/videos