Power CLI and Workflow manager for LLMs (core package)
Framework and no-code GUI for fine-tuning LLMs
A list of free LLM inference resources accessible via API
A high-throughput and memory-efficient inference and serving engine
A straightforward method for training your LLM
Enhances Tesseract OCR output using LLMs (local or API)
Accelerate local LLM inference and finetuning
Language Model Reinforcement Learning Environments frameworks
FlashInfer: Kernel Library for LLM Serving
Advanced LLM-powered brute-force tool combining AI intelligence
Run Local LLMs on Any Device. Open-source
LLM inference server with continuous batching & SSD caching
Let Claude (or any LLM) actually watch a video
lightweight package to simplify LLM API calls
Simple, Pythonic building blocks to evaluate LLM applications
Adding guardrails to large language models
Compress tool outputs, logs, files, and RAG chunks
Supercharge Your LLM Application Evaluations
Find the local LLM that actually runs and performs best
TokenSpeed is a speed-of-light LLM inference engine
State-of-the-art Parameter-Efficient Fine-Tuning
Framework to easily create LLM powered bots over any dataset
Open-source LLM Friendly Web Crawler & Scraper
A security scanner for custom LLM applications