Towards Efficient Self-Evolving Agent System
Unified KV Cache Compression Methods for Auto-Regressive Models
Learning to Reason with Search for LLMs via Reinforcement Learning
Cache-Augmented Generation: A Simple, Efficient Alternative to RAG
Scalable RL solution for advanced reasoning of language models
Official inference framework for 1-bit LLMs
Implementation of DeepLabCut
Designed for training LLM/VLM agents via RL
MobileLLM Optimizing Sub-billion Parameter Language Models
AI Agent Application Development Framework
AI assistant based on large models that can actively think and plan
MemU is an open-source memory framework for AI companions
Fast State-of-the-Art Static Embeddings
Production-grade platform for building agentic IM bots
Independent Auditing of AI Agents
Claude code for everything except coding
Pruna is a model optimization framework built for developers
World of apps for benchmarking interactive coding agent
Unified web UI for training and running open models locally
Superduper: Integrate AI models and machine learning workflows
Enterprise AI agent platform for workflows, models, and RAG apps
One-stop solution for creating your digital avatar from chat history
AI agents running research on single-GPU nanochat training
Apple Silicon (MLX) port of Karpathy's autoresearch
An agentic Machine Learning Engineer