Port of Facebook's LLaMA model in C/C++
Run models like Kimi-K2.5, GLM-5, DeepSeek, gpt-oss, Gemma, Qwen etc.
High-performance code intelligence MCP server
Port of OpenAI's Whisper model in C/C++
The Operator Splitting QP Solver
FAIR Sequence Modeling Toolkit 2
C++ and Python Examples
TEN, a voice agent framework to create conversational AI.
Open-source framework for conversational voice AI agents
AI video generator optimized for low VRAM and older GPUs use
Run GLM-5.2 (744B MoE) on a 25GB-RAM consumer machine
Run OpenClaw on a $5 chip
kaldi-asr/kaldi is the official location of the Kaldi project
Open-source vector similarity search for Postgres
Provides CTP stock options and Zhongtai Securities XTP
Low-latency AI inference engine optimized for mobile devices
Cloud-native open source data warehouse for analytics and AI queries
A fast image processing library with low memory needs
Your personal AI assistant at all-in 888KiB
A 2.78-trillion-parameter Kimi K3 running inference on a single CPU
Flux 2 image generation model pure C inference
MiniMax H3 inference engine for Mac computers
ESP32 desk dashboard that shows Claude Code usage
Run a 1-billion parameter LLM on a $10 board with 256MB RAM
Run the full 2.78-trillion-parameter Kimi K3 model