A high-throughput and memory-efficient inference and serving engine
Garnet is a remote cache-store from Microsoft Research
Web-based Traffic and Security Network Traffic Monitoring
Ultra quick message queue and streaming server
iperf3: A TCP, UDP, and SCTP network bandwidth measurement tool
A solid, high-performance, JDBC connection pool at last
The official Rust implementation of Conflux protocol
Open-source, scalable, and fault-tolerant MQTT broker
A high-performance inference system for large language models
Deep learning optimization library: makes distributed training easy
Parallax is a distributed model serving framework
Alibaba's high-performance LLM inference engine for diverse apps
Minimal Python framework for scalable AI inference servers fast
High-performance inference server for text embeddings models API layer
Running large language models on a single GPU
MiMo-V2-Flash: Efficient Reasoning, Coding, and Agentic Foundation
Distributed, in-memory key/value store and cache
AI memory OS for LLM and Agent systems
Efficient UDP unicast, UDP multicast, and IPC message transport
Node.js bindings for librdkafka
Go web framework benchmark
Low-latency REST API for serving text-embeddings
MII makes low-latency and high-throughput inference possible
Shardeum is an EVM based autoscaling blockchain
Postgresql R2DBC Driver