Port of Facebook's LLaMA model in C/C++
Efficient few-shot learning with Sentence Transformers
lightweight, standalone C++ inference engine for Google's Gemma models
Powering Amazon custom machine learning chips
Operating LLMs in production
AI interface for tinkerers (Ollama, Haystack RAG, Python)
Simplifies the local serving of AI models from any source
Replace OpenAI GPT with another LLM in your app
An Open-Source Programming Framework for Agentic AI
A Unified Library for Parameter-Efficient Learning
A unified framework for scalable computing
Serving system for machine learning models
PyTorch library of curated Transformer models and their components
LLMs and Machine Learning done easily
Sequence-to-sequence framework, focused on Neural Machine Translation