Powerful AI language model (MoE) optimized for efficiency/performance
Open-source, high-performance AI model with advanced reasoning
Open-source multi-speaker long-form text-to-speech model
Python inference and LoRA trainer package for the LTX-2 audio–video
Official Python inference and LoRA trainer package
Open Source Speech Language Model
Qwen3-TTS is an open-source series of TTS models
Models for object and human mesh reconstruction
Accurate × Fast × Comprehensive
AI cognitive-enhancement Skills based on Anthropic's J-space
A Powerful Native Multimodal Model for Image Generation
Netease Youdao's open-source embedding and reranker models
Video understanding codebase from FAIR for reproducing video models
Clean and efficient FP8 GEMM kernels with fine-grained scaling
Mixture-of-Experts Vision-Language Models for Advanced Multimodal
Unified Multimodal Understanding and Generation Models
MiniMax M2.1, a SOTA model for real-world dev & agents.
Analyze computation-communication overlap in V3/R1
A 0.1B Omni model trained from scratch
26m function call model that runs on incredibly small devices
Scaling Reinforcement Learning with LLMs
High-resolution models for human tasks
Instructions on how to use the Realtime API on Microcontrollers
Long-form streaming TTS system for multi-speaker dialogue generation
Open-source deep-learning framework