A Next-Generation Training Engine Built for Ultra-Large MoE Models
Based on AI Agent + MCP toolchain + penetration Skill orchestration
InvokeAI is a leading creative engine for Stable Diffusion models
950 line, minimal, extensible LLM inference engine built from scratch
RGBD video generation model conditioned on camera input
TokenSpeed is a speed-of-light LLM inference engine
Lightweight demo to build a conversational AI search engine quickly
Offline inference engine for art, real-time voice conversations
A community-supported supercharged version of paperless
Cloud-native open source data warehouse for analytics and AI queries
Pruna is a model optimization framework built for developers
An official Qdrant Model Context Protocol (MCP) server implementation
Comprehensive Gradio WebUI for audio processing
AI agent harness for AI coding agents
local-first semantic code search engine
Agentic IM Chatbot infrastructure
An Open Source package that allows video game creators
A modular graph-based Retrieval-Augmented Generation (RAG) system
A robust, efficient, low-latency speech-to-text library
A long-running autonomous coding agent powered by the Claude Agent
Generate audiobooks from EPUBs, PDFs and text with captions
Local long-term memory engine for AI apps with persistent storage
Low-latency AI inference engine optimized for mobile devices
GPU accelerated decision optimization
High-performance inference framework for large language models