TT-NN operator library, and TT-Metalium low level kernel programming
WebAssembly binding for llama.cpp - Enabling on-browser LLM inference
A high-performance inference engine for AI models
A.S.E (AICGSecEval) is a repository-level AI-generated code security
A course of learning LLM inference serving on Apple Silicon
TokenSpeed is a speed-of-light LLM inference engine
The official implementation of RAPTOR
Advanced LLM-powered brute-force tool combining AI intelligence
High-performance inference framework for large language models
Weaving the Digital Agent Galaxy
Mooncake is the serving platform for Kimi
AI-Powered Data Processing: Use LOTUS to process all of your datasets
Designed for text embedding and ranking tasks
How to optimize some algorithm in cuda
A guidance language for controlling large language models
Advanced language and coding AI model
State of the art LLM and coding model
An extensible framework for Personal Data Management
Build a modern LLM from scratch. Every line commented
Semi-Structured Agentic Framework. Workflows build themselves
Open-source enterprise-level AI knowledge base and MCP
Unified framework for building enterprise RAG pipelines
Scalable data pre processing and curation toolkit for LLMs
csghub-server is the backend server for CSGHub
GPT4V-level open-source multi-modal model based on Llama3-8B