Models for object and human mesh reconstruction
Video Object and Interaction Deletion
Qwen-Image is a powerful image generation foundation model
Qwen-Image-Layered: Layered Decomposition for Inherent Editablity
code for Mesh R-CNN, ICCV 2019
Uncommon Objects in 3D dataset
Codex plugin that turns attached object images into code-only
Recovering the Visual Space from Any Views
Qwen2.5-VL is the multimodal large language model series
Generating Immersive, Explorable, and Interactive 3D Worlds
A SOTA open-source image editing model
Official implementation of DreamCraft3D
General-purpose image editing model that delivers high-fidelity
NVIDIA Isaac GR00T N1.5 is the world's first open foundation model
Large Multimodal Models for Video Understanding and Editing
Tooling for the Common Objects In 3D dataset
Code release for "Masked-attention Mask Transformer
PyTorch implementation of YOLOv4
Learning Continuous Signed Distance Functions for Shape Representation
Code for "Image Generation from Scene Graphs", Johnson et al, CVPR 201