Structured outputs for llms
Python bindings for llama.cpp
Universal LLM Deployment Engine with ML Compilation
Run models like Kimi-K2.5, GLM-5, DeepSeek, gpt-oss, Gemma, Qwen etc.
Run Local LLMs on Any Device. Open-source
Emscripten: An LLVM-to-WebAssembly Compiler
Advanced LLM-powered brute-force tool combining AI intelligence
Fully automatic censorship removal for language models
An Agent Designed for Mathematical Modeling
A high-throughput and memory-efficient inference and serving engine
Port of Facebook's LLaMA model in C/C++
AirLLM 70B inference with single 4GB GPU
Easy-to-use LLM fine-tuning framework (LLaMA-2, BLOOM, Falcon
Build AI WhatsApp Bots with Pure Python
Language-model investigation agent with a terminal UI
Uses Qwen3-ASR, local LLM, Whisper, TEN-VAD
Scalable data pre processing and curation toolkit for LLMs
lightweight package to simplify LLM API calls
Powerful AI language model (MoE) optimized for efficiency/performance
Interact with your documents using the power of GPT
Korea Investment & Securities Open API Github
A guidance language for controlling large language models
Low-code app builder for RAG and multi-agent AI applications
Ongoing research training transformer models at scale
Open-source, high-performance AI model with advanced reasoning