Run Bonsai (1-bit) and Ternary-Bonsai language models locally
Awesome multilingual OCR toolkits based on PaddlePaddle
Official inference repo for FLUX.1 models
Official Python inference and LoRA trainer package
Programmatic access to the AlphaGenome model
AI PPT Track Terminator, the strongest PPT Skill ever
Lets make video diffusion practical
Tiny vision language model
Sharp Monocular Metric Depth in Less Than a Second
Code for running inference and finetuning with SAM 3 model
LTX-Video Support for ComfyUI
Text and image to video generation: CogVideoX and CogVideo
Inference script for Oasis 500M
Qwen3-TTS is an open-source series of TTS models
Reference PyTorch implementation and models for DINOv3
Python inference and LoRA trainer package for the LTX-2 audio–video
PyTorch code and models for the DINOv2 self-supervised learning
Diffusion Transformer with Fine-Grained Chinese Understanding
Pokee Deep Research Model Open Source Repo
GLM-4.5: Open-source LLM for intelligent agents by Z.ai
Models for object and human mesh reconstruction
Qwen3-Coder is the code version of Qwen3
Generate Any 3D Scene in Seconds
Recovering the Visual Space from Any Views
Industrial-level controllable zero-shot text-to-speech system