Open-source, high-performance AI model with advanced reasoning
Official repository for LTX-Video
Open-source large language model family from Tencent Hunyuan
Native and Compact Structured Latents for 3D Generation
Robust Speech Recognition Across Languages, Dialects
Ultra-Efficient LLMs on End Device
Designed for text embedding and ranking tasks
Open-source industrial-grade ASR models
Diversity-driven optimization and large-model reasoning ability
Powerful AI language model (MoE) optimized for efficiency/performance
Fast stable diffusion on CPU and AI PC
Ling-V2 is a MoE LLM provided and open-sourced by InclusionAI
Advanced language and coding AI model
Python inference and LoRA trainer package for the LTX-2 audio–video
Reference PyTorch implementation and models for DINOv3
Qwen3-TTS is an open-source series of TTS models
High-Resolution Image Synthesis with Latent Diffusion Models
GLM-4.5: Open-source LLM for intelligent agents by Z.ai
Miso TTS is an 8 billion, highly emotive text-to-speech model
Repo of Qwen2-Audio chat & pretrained large audio language model
A GPT-4o Level MLLM for Vision, Speech and Multimodal Live Streaming
Code for running inference and finetuning with SAM 3 model
Hunyuan Translation Model Version 1.5
Achieving 3+ generation speedup on reasoning tasks
Open-source image generative foundation model