Qwen3-TTS is an open-source series of TTS models
26m function call model that runs on incredibly small devices
Official inference repo for FLUX.1 models
DeepSeek Coder: Let the Code Write Itself
Infinite Worlds with Versatile Interactions
A theoretical reconstruction of the Claude Mythos architecture
A Family of Open Sourced Music Foundation Models
Text and image to video generation: CogVideoX and CogVideo
Convert Google Gemini web into OpenAI-compatible API
Project Lyra: Open Generative 3D World Models
Fast-stable-diffusion + DreamBooth
Netease Youdao's open-source embedding and reranker models
Advancing Open-source World Models
Hunyuan Translation Model Version 1.5
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
Tiny vision language model
High-Resolution Image Synthesis with Latent Diffusion Models
Audio Language Models are Few-Shot Learners
A 0.1B Omni model trained from scratch
MOSS‑TTS Family open‑source speech and sound generation model
State-of-the-art Image & Video CLIP, Multimodal Large Language Models
Inference script for Oasis 500M
Generate Any 3D Scene in Seconds
Unified Multimodal Understanding and Generation Models
OCR expert VLM powered by Hunyuan's native multimodal architecture