Powerful AI language model (MoE) optimized for efficiency/performance
An easy-to-understand framework for LLM samplers
Easy token price estimates for 400+ LLMs. TokenOps
Persistent context and multi-instance coordination
This repository contains the official implementation of FastVLM
Token-Efficient AI Agent with same budget, higher intelligence density
Compress tool outputs, logs, files, and RAG chunks
14-stage Fusion Pipeline for LLM token compression
Bidirectional token-classification model for identifiable info
An open source Flask extension that provides JWT support
Real-time Claude Code usage monitor with predictions and warnings
Simple integration of Flask and WTForms, including CSRF
Complete Two-Factor Authentication for Django
LLM-based Reinforcement Learning audio edit model
Google Mail registration
Why use many token when few token do trick
Open-source, high-performance AI model with advanced reasoning
The best way to use Hermes Agent from the web or from your phone
SOTA on-device LLMs, small yet powerful
Create prompt-friendly codebase digests from any Git repository URL
Real-time multi-AI collaboration: Claude, Codex & Gemini
The official CLI to interact with Kaggle
Block Diffusion for Ultra-Fast Speculative Decoding
Claude Engineer is an interactive command-line interface (CLI)
A Powerful Native Multimodal Model for Image Generation