Fast, flexible LLM inference
Jlama is a modern LLM inference engine for Java
LLM abstractions that aren't obstructions
Production ready toolkit to run AI locally
Build a modern LLM from scratch. Every line commented
One beautiful Ruby API for OpenAI, Anthropic, Gemini, Bedrock
Accelerate local LLM inference and finetuning
Chat with your documents using local AI
Fully private LLM chatbot that runs entirely with a browser
A large model training tool that supports training large models