Python-free Rust inference server
High-performance inference server for text embeddings models API layer
Fast and efficient unstructured data extraction
Switchyard lets LLM applications route traffic across models
Ghost in your shell. Ante is a self-contained agent harness
Fastest LLM inference runtime for Apple Silicon
Fast ML inference & training for ONNX models in Rust
Package and deploy machine learning models using Docker containers
Command-line tool for Drive, Gmail, Calendar, Sheets, Docs, Chat, etc.
Rust async runtime based on io-uring