A full-stack codebase for training and evaluating speculative decoding
A HEVC/H.265 Web Player
Fountain-coded QR file transfer
Runtime type system for IO decoding/encoding
Redundancy-aware KV Cache Compression for Reasoning Models
Block Diffusion for Ultra-Fast Speculative Decoding
FlashMLA: Efficient Multi-head Latent Attention Kernels
Jessibuca is an open source pure H5 live streaming player
Runtime type system for IO decoding/encoding
Clojure JSON and JSON SMILE (binary json format) encoding/decoding
Achieving 3+ generation speedup on reasoning tasks
Java MQTT lightweight broker
Fast LLM speculative inference server for consumer hardware
Multiplatform C++ library for parsing and crafting of network packets
A Swift wrapper for the FFmpeg API
The “Quite OK Image Format” for fast, lossless image compression
Pythonic bindings for FFmpeg's libraries
A simple, yet elegant, HTTP library.
The RF and reverse engineering framework for everyone
Fast Multimodal LLM on Mobile Devices
High-performance Inference and Deployment Toolkit for LLMs and VLMs
Document Image Parsing via Heterogeneous Anchor Prompting”
tiktoken is a fast BPE tokeniser for use with OpenAI's models
Official repository for LTX-Video
Library to encode and decode images in WebP format