AirLLM 70B inference with single 4GB GPU
Personal Information “Leakage ” Detection Interface
HunyuanVideo: A Systematic Framework For Large Video Generation Model
The best way to use Hermes Agent from the web or from your phone
Building an Intelligent Agent from Scratch
A personal AI assistant, easy to install
MemU is an open-source memory framework for AI companions
Open-source large language model family from Tencent Hunyuan
Unified KV Cache Compression Methods for Auto-Regressive Models
A lightweight, powerful framework for multi-agent workflows
Neural Network architecture based on ideas of the original LSTM
A visual, example-driven guide to Claude Code
Simple package for monitoring and control your NVIDIA Jetson
A Python library for audio
Accessible large language models via k-bit quantization for PyTorch
Redundancy-aware KV Cache Compression for Reasoning Models
Low-latency AI inference engine optimized for mobile devices
AI Agent Source Code Deep Research Report
Practice made claude perfect
Windrecorder is a memory search app by records everything
DeepEP: an efficient expert-parallel communication library
Unified web UI for training and running open models locally
Running large language models on a single GPU
A Full-Link Guide to Using Codex from Installation to Real-World Cases
Persistent context and multi-instance coordination