Z80-μLM is a 2-bit quantized language model
A command-line productivity tool powered by AI large language models
Dev tools, env vars, task runner
FlashInfer: Kernel Library for LLM Serving
Vector Database for the next generation of AI applications
Static type checker for Python
The official Meta Llama 3 GitHub site
DeepSeek Coder: Let the Code Write Itself
Composable building blocks to build Llama Apps
An elegent pytorch implement of transformers
Simple PDF generation for Python
Big Model Application Development Practice 1
950 line, minimal, extensible LLM inference engine built from scratch
High-performance fake data generator for Python
Generate music based on natural language prompts using LLMs
Fast Python library for SEGY files
Speech-to-text, text-to-speech, and speaker recognition
Leaderboard Comparing LLM Performance at Producing Hallucinations
Examples and tutorials to help developers build AI systems
StarVector is a foundation model for SVG generation
Weaving the Digital Agent Galaxy
Director, Screenwriter, Producer, and Video Generator All-in-One
Train a 26M-parameter GPT from scratch in just 2h
File Parser optimised for LLM Ingestion with no loss
INT4/INT5/INT8 and FP16 inference on CPU for RWKV language model