LM Studio Apple MLX engine
TTS with kokoro and onnx runtime
Fastest LLM inference runtime for Apple Silicon
MCP server for interfacing with Godot game engine
DeepSeek 4 Flash local inference engine for Metal
SQL-Driven RAG Engine
Universal LLM Deployment Engine with ML Compilation
950 line, minimal, extensible LLM inference engine built from scratch
Multi-Agent daTa geneRation Infra and eXperimentation framework
Run a 1-billion parameter LLM on a $10 board with 256MB RAM
Fast Multimodal LLM on Mobile Devices
Smart LLM router
NestJS Helper + AI Chatbot Development
Next-generation AI Agent Optimization Platform