Run Bonsai (1-bit) and Ternary-Bonsai language models locally
Official inference repo for FLUX.1 models
GLM-5: From Vibe Coding to Agentic Engineering
Fast-stable-diffusion + DreamBooth
Your clothes, extracted and organized with gpt-image
The most powerful local music generation model
Genome modeling and design across all domains of life
Official Python inference and LoRA trainer package
Convert Google Gemini web into OpenAI-compatible API
One-click generation of various gameplay, no prompt words required
Gemma 4 26B-A4B inference in ~2 GB of RAM on any M-series MacBook
MiniMax M2.1, a SOTA model for real-world dev & agents.
Python SDK for Claude Agent
Open-source, high-performance AI model with advanced reasoning
Contexts Optical Compression
26m function call model that runs on incredibly small devices
AlphaFold 3 inference pipeline
C#/.NET binding of llama.cpp, including LLaMa/GPT model inference
Official inference repo for FLUX.2 models
GLM-4 series: Open Multilingual Multimodal Chat LMs
Qwen3 is the large language model series developed by Qwen team
HY-Motion model for 3D character animation generation
Fast, Sharp & Reliable Agentic Intelligence
Qwen3.6 is the large language model series developed by Qwen team
Miso TTS is an 8 billion, highly emotive text-to-speech model