Official inference repo for FLUX.1 models
Qwen3-TTS is an open-source series of TTS models
Fast-stable-diffusion + DreamBooth
Text and image to video generation: CogVideoX and CogVideo
Infinite Worlds with Versatile Interactions
DeepSeek Coder: Let the Code Write Itself
26m function call model that runs on incredibly small devices
Tiny vision language model
A 0.1B Omni model trained from scratch
A Family of Open Sourced Music Foundation Models
Project Lyra: Open Generative 3D World Models
A theoretical reconstruction of the Claude Mythos architecture
Recovering the Visual Space from Any Views
Unified Multimodal Understanding and Generation Models
Convert Google Gemini web into OpenAI-compatible API
Netease Youdao's open-source embedding and reranker models
High-Resolution Image Synthesis with Latent Diffusion Models
Inference script for Oasis 500M
Audio Language Models are Few-Shot Learners
High-resolution models for human tasks
Official code base for LeWorldModel: Stable End-to-End Joint-Embedding
Advancing Open-source World Models
OCR expert VLM powered by Hunyuan's native multimodal architecture
Open image model at the forefront of design
Open-source image generative foundation model