LLM inference server with continuous batching & SSD caching
Self-hosted game stream host for Moonlight
An ASGI web server, for Python
gpt-oss-120b and gpt-oss-20b are two open-weight language models
ChatGLM3 series: Open Bilingual Chat LLMs | Open Source Bilingual Chat
Web apps in pure Python
Seamlessly extend your preferred base images to be Lambda compatible
LLM plugin providing access to models running on an Ollama server
openvpn-monitor is a web based OpenVPN monitor
Matrix reference homeserver
The premiere source of truth powering network automation
Python binding to the Apache Tika™ REST services
Googles NotebookLM but local
Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model
A full-stack AI Red Teaming platform securing AI ecosystems
OCR model for complex documents with layout-aware structured outputs
Multi-LoRA inference server that scales to 1000s of fine-tuned LLMs
Run Bonsai (1-bit) and Ternary-Bonsai language models locally
The fastest way to build data apps in Python
Unified web UI for training and running open models locally
Open-source LLM Friendly Web Crawler & Scraper
Define and run multi-container applications with Docker
Supercharge Your LLM with the Fastest KV Cache Layer
Lemonade helps users run local LLMs with the highest performance
A persistent workspace for development work that self-improves