Python bindings for llama.cpp
Designed for text embedding and ranking tasks
Generates original ARC-AGI-1-style tasks distribution-matched
Reference PyTorch implementation and models for DINOv3
LTX-Video Support for ComfyUI
Open-source, high-performance AI model with advanced reasoning
Qwen3 is the large language model series developed by Qwen team
4M: Massively Multimodal Masked Modeling
Research code artifacts for Code World Model (CWM)
Achieving 3+ generation speedup on reasoning tasks
Qwen3-Coder is the code version of Qwen3
Mixture-of-Experts Vision-Language Models for Advanced Multimodal
Text and image to video generation: CogVideoX and CogVideo
Open-source large language model family from Tencent Hunyuan
Phi-3.5 for Mac: Locally-run Vision and Language Models
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
The most powerful local music generation model
High-resolution models for human tasks
Qwen-Image-Layered: Layered Decomposition for Inherent Editablity
MapAnything: Universal Feed-Forward Metric 3D Reconstruction
Wan2.2: Open and Advanced Large-Scale Video Generative Model
Lets make video diffusion practical
Wan2.1: Open and Advanced Large-Scale Video Generative Model
Diversity-driven optimization and large-model reasoning ability
Audio Language Models are Few-Shot Learners