Multimodal Diffusion with Representation Alignment
PyTorch implementation of JiT
Python inference and LoRA trainer package for the LTX-2 audio–video
Controllable & emotion-expressive zero-shot TTS
Industrial-level controllable zero-shot text-to-speech system
DeepSeek Coder: Let the Code Write Itself
Qwen3.6 is the large language model series developed by Qwen team
A Powerful Native Multimodal Model for Image Generation
Towards self-verifiable mathematical reasoning
Clean and efficient FP8 GEMM kernels with fine-grained scaling
Advanced language and coding AI model
A 0.1B Omni model trained from scratch
OpenTinker is an RL-as-a-Service infrastructure for foundation models
Instructions on how to use the Realtime API on Microcontrollers
tiktoken is a fast BPE tokeniser for use with OpenAI's models
Provides convenient access to the Anthropic REST API from any Python 3
MiniMax-M2, a model built for Max coding & agentic workflows
Qwen2.5-Coder is the code version of Qwen2.5, the large language model
Self-evolving AI model for agents, coding, and complex workflows
OpenAI’s open-weight 120B model optimized for reasoning and tooling
Flagship MoE model for advanced reasoning, coding, and agents
High-compute ultra-reasoning model surpassing model surpassing GPT-5
Dense multimodal Qwen model for coding, agents, and long context
Open multimodal model for coding, agents, and long-context tasks
Omnimodal AI model for agents, coding, and long-context tasks