Powerful AI language model (MoE) optimized for efficiency/performance
Token-Efficient AI Agent with same budget, higher intelligence density
This repository contains the official implementation of FastVLM
Bidirectional token-classification model for identifiable info
Easy token price estimates for 400+ LLMs. TokenOps
LLM-based Reinforcement Learning audio edit model
Why use many token when few token do trick
Persistent context and multi-instance coordination
Implementation of Phenaki Video, which uses Mask GIT
Compress tool outputs, logs, files, and RAG chunks
Open-source, high-performance AI model with advanced reasoning
14-stage Fusion Pipeline for LLM token compression
Real-time Claude Code usage monitor with predictions and warnings
The best way to use Hermes Agent from the web or from your phone
A Powerful Native Multimodal Model for Image Generation
Offical Implementation for "Recursive Multi-Agent Systems"
OpenSpace: Make Your Agents: Smarter, Low-Cost, Self-Evolving
Real-time multi-AI collaboration: Claude, Codex & Gemini
Minimal reproduction of OneRec
MoBA: Mixture of Block Attention for Long-Context LLMs
Block Diffusion for Ultra-Fast Speculative Decoding
Context management for Claude Code. Hooks maintain state via ledgers
Create prompt-friendly codebase digests from any Git repository URL
AWS Skills for Agents
tiktoken is a fast BPE tokeniser for use with OpenAI's models