Hackable and optimized Transformers building blocks
Open-source large language model family from Tencent Hunyuan
High-Resolution Image Synthesis with Latent Diffusion Models
Stable Diffusion WebUI Forge is a platform on top of Stable Diffusion
The most powerful local music generation model
Lets make video diffusion practical
MapAnything: Universal Feed-Forward Metric 3D Reconstruction
Text and image to video generation: CogVideoX and CogVideo
Ring is a reasoning MoE LLM provided and open-sourced by InclusionAI
The official repo of Qwen chat & pretrained large language model
Memory-efficient and performant finetuning of Mistral's models
An experimental version of DeepSeek model
A Multi-Modal World Model for Reconstructing, Generating, Simulation
PyTorch code and models for the DINOv2 self-supervised learning
tiktoken is a fast BPE tokeniser for use with OpenAI's models
Tencent Hunyuan Multimodal diffusion transformer (MM-DiT) model
gpt-oss-120b and gpt-oss-20b are two open-weight language models
ChatGLM-6B: An Open Bilingual Dialogue Language Model
Z80-μLM is a 2-bit quantized language model
Advancing Open-source World Models
GPT4V-level open-source multi-modal model based on Llama3-8B
Diversity-driven optimization and large-model reasoning ability
Chinese and English multimodal conversational language model
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
A trainable PyTorch reproduction of AlphaFold 3