Powerful AI language model (MoE) optimized for efficiency/performance
Easy token price estimates for 400+ LLMs. TokenOps
Persistent context and multi-instance coordination
Token-Efficient AI Agent with same budget, higher intelligence density
This repository contains the official implementation of FastVLM
Compress tool outputs, logs, files, and RAG chunks
14-stage Fusion Pipeline for LLM token compression
Bidirectional token-classification model for identifiable info
Real-time Claude Code usage monitor with predictions and warnings
LLM-based Reinforcement Learning audio edit model
Why use many token when few token do trick
Open-source, high-performance AI model with advanced reasoning
The best way to use Hermes Agent from the web or from your phone
SOTA on-device LLMs, small yet powerful
Create prompt-friendly codebase digests from any Git repository URL
Real-time multi-AI collaboration: Claude, Codex & Gemini
Block Diffusion for Ultra-Fast Speculative Decoding
A Powerful Native Multimodal Model for Image Generation
tiktoken is a fast BPE tokeniser for use with OpenAI's models
OpenSpace: Make Your Agents: Smarter, Low-Cost, Self-Evolving
Context management for Claude Code. Hooks maintain state via ledgers
Large Language Model Text Generation Inference
Provides line-oriented text file editing capabilities
Build your own Cowork, AI Scientist and other SoTA Agents
Offical Implementation for "Recursive Multi-Agent Systems"