Structured outputs for llms
Python bindings for llama.cpp
Run Local LLMs on Any Device. Open-source
Fully automatic censorship removal for language models
Advanced LLM-powered brute-force tool combining AI intelligence
LLM inference server with continuous batching & SSD caching
An Agent Designed for Mathematical Modeling
A high-throughput and memory-efficient inference and serving engine
AirLLM 70B inference with single 4GB GPU
Build AI WhatsApp Bots with Pure Python
Easy-to-use LLM fine-tuning framework (LLaMA-2, BLOOM, Falcon
Language-model investigation agent with a terminal UI
Powerful AI language model (MoE) optimized for efficiency/performance
Scalable data pre processing and curation toolkit for LLMs
Open-source, high-performance AI model with advanced reasoning
lightweight package to simplify LLM API calls
Korea Investment & Securities Open API Github
Uses Qwen3-ASR, local LLM, Whisper, TEN-VAD
A guidance language for controlling large language models
The official Meta Llama 3 GitHub site
Interact with your documents using the power of GPT
Compress tool outputs, logs, files, and RAG chunks
Universal LLM Deployment Engine with ML Compilation
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
Ongoing research training transformer models at scale