TencentDB Agent Memory delivers fully local long-term memory for AI
AirLLM 70B inference with single 4GB GPU
Building an Intelligent Agent from Scratch
Unified KV Cache Compression Methods for Auto-Regressive Models
Neural Network architecture based on ideas of the original LSTM
MemU is an open-source memory framework for AI companions
Real-time NVIDIA GPU dashboard
Open-source large language model family from Tencent Hunyuan
Demo of a customer service use case implemented with the OpenAI Agents
Redundancy-aware KV Cache Compression for Reasoning Models
AI Agent Source Code Deep Research Report
Run a 1-billion parameter LLM on a $10 board with 256MB RAM
Running large language models on a single GPU
Persistent context and multi-instance coordination
The repository provides code for running inference with SAM 2
Self-evolving autonomous agent framework
A step-by-step guide to build your own AI agent
A Web UI for easy subtitle using whisper model
Memory-efficient and performant finetuning of Mistral's models
14-stage Fusion Pipeline for LLM token compression
Designed for training LLM/VLM agents via RL
Official plugin for OpenClaw that exports agent traces to Opik
Python-free Rust inference server
LangChain4j is an open-source Java library
A high-quality rapid TTS voice cloning model