Official inference repo for FLUX.2 models
A Customizable Image-to-Video Model based on HunyuanVideo
High-Resolution Image Synthesis with Latent Diffusion Models
A Family of Open Sourced Music Foundation Models
Tencent Hunyuan Multimodal diffusion transformer (MM-DiT) model
Personalize Any Characters with a Scalable Diffusion Transformer
gpt-oss-120b and gpt-oss-20b are two open-weight language models
Z80-μLM is a 2-bit quantized language model
Reference PyTorch implementation and models for DINOv3
A SOTA open-source image editing model
Industrial-level controllable zero-shot text-to-speech system
1B text generation model based on the HRM architecture
Inference script for Oasis 500M
Generate Any 3D Scene in Seconds
The official PyTorch implementation of Google's Gemma models
VMZ: Model Zoo for Video Modeling
Fast and Universal 3D reconstruction model for versatile tasks
High-Fidelity and Controllable Generation of Textured 3D Assets
Official DeiT repository
Dataset of GPT-2 outputs for research in detection, biases, and more
Official code for Style Aligned Image Generation via Shared Attention
This repository contains the official implementation of research
Towards Robust Blind Face Restoration with Codebook Lookup Transformer
Repo for external large-scale work
A latent text-to-image diffusion model