Wan2.2: Open and Advanced Large-Scale Video Generative Model
Multimodal Diffusion with Representation Alignment
A Pragmatic VLA Foundation Model
PyTorch implementation of JiT
Python inference and LoRA trainer package for the LTX-2 audio–video
A trainable PyTorch reproduction of AlphaFold 3
Advancing Open-source World Models
Controllable & emotion-expressive zero-shot TTS
LLM-based Reinforcement Learning audio edit model
Open image model at the forefront of design
Industrial-level controllable zero-shot text-to-speech system
HY-Motion model for 3D character animation generation
DeepSeek Coder: Let the Code Write Itself
Designed for text embedding and ranking tasks
A GPT-4o Level MLLM for Vision, Speech and Multimodal Live Streaming
Qwen3.6 is the large language model series developed by Qwen team
Towards self-verifiable mathematical reasoning
A Powerful Native Multimodal Model for Image Generation
Clean and efficient FP8 GEMM kernels with fine-grained scaling
Advanced language and coding AI model
A 0.1B Omni model trained from scratch
OpenTinker is an RL-as-a-Service infrastructure for foundation models
Instructions on how to use the Realtime API on Microcontrollers
State of the art LLM and coding model
tiktoken is a fast BPE tokeniser for use with OpenAI's models