BaseRT is a local large language model inference runtime optimized for Apple Silicon computers. It accelerates model execution through hand-written Metal kernels and requires an M1 or newer Mac running macOS 14 or later. A unified command-line interface can download models from Hugging Face, convert checkpoints, launch chats, benchmark performance, and inspect model packages. Its server implements OpenAI-compatible chat, completion, embedding, transcription, tool-calling, and multimodal endpoints. The custom .base format supports affine quantization from Q2 through Q8, optional AWQ calibration, and signed model bundles. Stable C interfaces connect the engine with Python, Node.js, Rust, and Swift applications. The repository contains the open CLI, format specifications, bindings, documentation, and benchmarks, while the prebuilt inference engine uses a separate license.

Features

  • Metal-accelerated Apple Silicon inference
  • Unified model management CLI
  • OpenAI-compatible local API server
  • Text, image, and audio model support
  • Q2 through Q8 model quantization
  • Python, Node.js, Rust, and Swift bindings

Project Samples

Project Activity

See All Activity >

License

Apache License V2.0

Follow BaseRT

BaseRT Web Site

Other Useful Business Software
Our Free Plans just got better! | Auth0 Icon
Our Free Plans just got better! | Auth0

With up to 25k MAUs and unlimited Okta connections, our Free Plan lets you focus on what you do best—building great apps.

You asked, we delivered! Auth0 is excited to expand our Free and Paid plans to include more options so you can focus on building, deploying, and scaling applications without having to worry about your security. Auth0 now, thank yourself later.
Try free now
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of BaseRT!

Additional Project Details

Programming Language

Rust

Related Categories

Rust Artificial Intelligence Software

Registered

6 days ago