A collection of scientific methods, processes, algorithms
Netflix’s Workflow Orchestrator
Faster and easier training and deployments
Running large language models on a single GPU
Instant neural graphics primitives: lightning fast NeRF and more
Designed for training LLM/VLM agents via RL
Agent framework that enables tool-use agent tasks
A @ClickHouse fork that supports high-performance vector search
Search + Chat = SearChat(AI Chat with Search)
Alibaba's high-performance LLM inference engine for diverse apps
Streamlines and simplifies prompt design for both developers
Hypernetworks that adapt LLMs for specific benchmark tasks
Run a 1-billion parameter LLM on a $10 board with 256MB RAM
Demystify RAG by building it from scratch
UCCL is an efficient communication library for GPUs
Cloud-native runtime for agentic AI
Your Personal Research Multi-Tool
Collect, organize, use, and share, all in OmniBox
Learning to Reason with Search for LLMs via Reinforcement Learning
A high-performance inference engine for AI models
Korvus is a search SDK that unifies the entire RAG pipeline
Domain analysis toolkit for DNS, IP, and WHOIS lookups
Cache-Augmented Generation: A Simple, Efficient Alternative to RAG
270+ Claude Code plugins with 739 agent skills
A tension reasoning engine over 131 S-class problems