TencentDB Agent Memory delivers fully local long-term memory for AI
Port of OpenAI's Whisper model in C/C++
AirLLM 70B inference with single 4GB GPU
Personal Information “Leakage ” Detection Interface
Building an Intelligent Agent from Scratch
A stream multiplexing library for Golang with minimal memory usage
Provides a way to profile code
Lightweight Java library developed by Alibaba for reading and writing
Library for the numerical simulation of closed as well as open quantum
Unified KV Cache Compression Methods for Auto-Regressive Models
Neural Network architecture based on ideas of the original LSTM
Simple package for monitoring and control your NVIDIA Jetson
MemU is an open-source memory framework for AI companions
Real-time NVIDIA GPU dashboard
Demo of a customer service use case implemented with the OpenAI Agents
Open-source large language model family from Tencent Hunyuan
Fast and Lightweight Logs and Metrics processor for Linux, BSD, OSX
Redundancy-aware KV Cache Compression for Reasoning Models
AI Agent Source Code Deep Research Report
FilterIterator implementation that filters files
VisualVM is an All-in-One Java Troubleshooting Tool
Run a 1-billion parameter LLM on a $10 board with 256MB RAM
DeepEP: an efficient expert-parallel communication library
waifu2x converter ncnn version, run fast GPU with vulkan
Running large language models on a single GPU