Python-free Rust inference server
Codex Switch & Instruct desktop manager
Fast, local-first web content extraction for LLMs
Fast and efficient unstructured data extraction
Fastest LLM inference runtime for Apple Silicon
Fast, flexible LLM inference
Ghost in your shell. Ante is a self-contained agent harness
Rust async runtime based on io-uring