Fast, small, and fully autonomous AI assistant infrastructure
A personal AI assistant that evolves with you
Run a 1-billion parameter LLM on a $10 board with 256MB RAM
Low-latency AI inference engine optimized for mobile devices
Agent framework and applications built upon Qwen>=3.0
Running large language models on a single GPU
Unified web UI for training and running open models locally
Persistent context and multi-instance coordination
A Python library for audio
Developer friendly Natural Language Processing
ReFT: Representation Finetuning for Language Models
The repository provides code for running inference with SAM 2
Self-evolving autonomous agent framework
High-speed Large Language Model Serving for Local Deployment
One brain, many harnesses. Portable .agent/ folder
A comprehensive collection of Agent Skills for context engineering
Drag & drop UI to build your customized LLM flow
Gradient boosting framework based on decision tree algorithms
LLM inference in C/C++
A step-by-step guide to build your own AI agent
Benchmarking synthetic data generation methods
A Web UI for easy subtitle using whisper model
Memory-efficient and performant finetuning of Mistral's models
LLM training in simple, raw C/CUDA
Turns Data and AI algorithms into production-ready web applications