High-Resolution 3D Assets Generation with Large Scale Diffusion Models
Official inference repo for FLUX.2 models
Official Python inference and LoRA trainer package
Recovering the Visual Space from Any Views
MiniMax H3 is a general-purpose, omni-modal generative system
High-Resolution Image Synthesis with Latent Diffusion Models
This repository contains the official implementation of FastVLM
Repo for SeedVR2 & SeedVR
Qwen2.5-VL is the multimodal large language model series
A Customizable Image-to-Video Model based on HunyuanVideo
Wan2.1: Open and Advanced Large-Scale Video Generative Model
Native and Compact Structured Latents for 3D Generation
Official repository for LTX-Video
GPT4V-level open-source multi-modal model based on Llama3-8B
Reference PyTorch implementation and models for DINOv3
Text and image to video generation: CogVideoX and CogVideo
Genome modeling and design across all domains of life
Sharp Monocular Metric Depth in Less Than a Second
Open image model at the forefront of design
Programmatic access to the AlphaGenome model
Implementation of the Surya Foundation Model for Heliophysics
High-resolution models for human tasks
A Unified Framework for Text-to-3D and Image-to-3D Generation
Generate Any 3D Scene in Seconds
Global weather forecasting model using graph neural networks and JAX