Make any agent harness multimodal-native
Running a 28.9M parameter LLM on an $8 microcontroller
Large Language Model Text Generation Inference
A solution to build and deploy MCP agents and applications
A Frontier Mathematical Coding Agent
The fast, Pythonic way to build Model Context Protocol servers
Visualizer for neural network, deep learning, machine learning models
local-first semantic code search engine
InvokeAI is a leading creative engine for Stable Diffusion models
The ultimate RAG for your monorepo
Open-source tool designed to enhance the efficiency of workloads
TokenSpeed is a speed-of-light LLM inference engine
Expose your FastAPI endpoints as Model Context Protocol (MCP) tools
A self-hosted open source photo management service
Official inference framework for 1-bit LLMs
Convert codebases into structured prompts optimized for LLM analysis
A proof-of-concept jupyter extension which converts english queries
Visualization language that lets AI agents create expressive charts
Specification and documentation for the Model Context Protocol
Jev-like family of decision models built on top of Qwen3.5/3.8
Centralized agent control plane for governing runtime agent behavior
🐈 nanobot: The Ultra-Lightweight Clawdbot / OpenClaw
This repository contains the official implementation of research
Low-latency REST API for serving text-embeddings
Standardized Serverless ML Inference Platform on Kubernetes