Official inference repo for FLUX.1 models
Qwen3-TTS is an open-source series of TTS models
Convert Google Gemini web into OpenAI-compatible API
Z80-μLM is a 2-bit quantized language model
A 0.1B Omni model trained from scratch
26m function call model that runs on incredibly small devices
A Family of Open Sourced Music Foundation Models
Text and image to video generation: CogVideoX and CogVideo
DeepSeek Coder: Let the Code Write Itself
A theoretical reconstruction of the Claude Mythos architecture
Project Lyra: Open Generative 3D World Models
Infinite Worlds with Versatile Interactions
Fast-stable-diffusion + DreamBooth
Netease Youdao's open-source embedding and reranker models
Recovering the Visual Space from Any Views
Open-source image generative foundation model
Tiny vision language model
High-Resolution Image Synthesis with Latent Diffusion Models
Generate Any 3D Scene in Seconds
scikit-learn compatible tabular foundation model
High-resolution models for human tasks
Unified Multimodal Understanding and Generation Models
Audio Language Models are Few-Shot Learners
A Systematic Framework for Interactive World Modeling
OCR expert VLM powered by Hunyuan's native multimodal architecture