A frontier, first-principles handbook
Make your agents learn from experience
Central interface to connect your LLM's with external data
Compress tool outputs, logs, files, and RAG chunks
Open-source, high-performance AI model with advanced reasoning
MoBA: Mixture of Block Attention for Long-Context LLMs
Agentic, Reasoning, and Coding (ARC) foundation models
Cache-Augmented Generation: A Simple, Efficient Alternative to RAG
The official implementation of RAPTOR
LongBench v2 and LongBench (ACL 25'&24')
Large-language-model & vision-language-model based on Linear Attention
MemoryOS is designed to provide a memory operating system
SimpleMem: Efficient Lifelong Memory for LLM Agents
Qwen3-Coder is the code version of Qwen3
Qwen3 is the large language model series developed by Qwen team
The official repo of Qwen chat & pretrained large language model
the terminal client for Ollama
Open-weight, large-scale hybrid-attention reasoning model
AI-powered penetration testing assistant using local LLM on linux
Advanced techniques for RAG systems
Claude + Obsidian knowledge companion
Open-source large language model family from Tencent Hunyuan
ChatGLM3 series: Open Bilingual Chat LLMs | Open Source Bilingual Chat
GLM-4 series: Open Multilingual Multimodal Chat LMs
A Next-Generation Training Engine Built for Ultra-Large MoE Models