Run Bonsai (1-bit) and Ternary-Bonsai language models locally
A modern desktop solution built for the DeepSeek Harness (DSH) plugin
Official inference repo for FLUX.1 models
GLM-5: From Vibe Coding to Agentic Engineering
One-click local MCP server installation in desktop apps
Fast-stable-diffusion + DreamBooth
Your clothes, extracted and organized with gpt-image
Diffusion Bee is the easiest way to run Stable Diffusion locally
Genome modeling and design across all domains of life
One-click generation of various gameplay, no prompt words required
Convert Google Gemini web into OpenAI-compatible API
The most powerful local music generation model
Official Python inference and LoRA trainer package
Gemma 4 26B-A4B inference in ~2 GB of RAM on any M-series MacBook
Python SDK for Claude Agent
MiniMax M2.1, a SOTA model for real-world dev & agents.
Contexts Optical Compression
Open-source, high-performance AI model with advanced reasoning
26m function call model that runs on incredibly small devices
AlphaFold 3 inference pipeline
C#/.NET binding of llama.cpp, including LLaMa/GPT model inference
Official inference repo for FLUX.2 models
Qwen3 is the large language model series developed by Qwen team
GLM-4 series: Open Multilingual Multimodal Chat LMs
HY-Motion model for 3D character animation generation