Running large language models on a single GPU
Persistent context and multi-instance coordination
The repository provides code for running inference with SAM 2
JetCache is a Java cache framework
Retro emulation for the ODROID-GO and other ESP32 devices
Self-evolving autonomous agent framework
A fast cache that automatically deletes the least recently used items
Frame profiler
A process for exposing JMX Beans via HTTP for Prometheus consumption
LangChain for Go, the easiest way to write LLM-based programs in Go
C++ learning and interview guide aimed at back-end systems developers
A step-by-step guide to build your own AI agent
A Web UI for easy subtitle using whisper model
The basic implementation for chunk upload with multiple providers
Memory-efficient and performant finetuning of Mistral's models
Swift Featured Projects in brain Mapping
Allows exporting any serializable PHP data structure to plain PHP code
14-stage Fusion Pipeline for LLM token compression
Designed for training LLM/VLM agents via RL
A suite of Monitoring Plugins (formerly known as nagios-plugins)
Official plugin for OpenClaw that exports agent traces to Opik
Random version of microemacs with my private modificatons
Python-free Rust inference server
Streaming downloads using Net::HTTP, http.rb or HTTPX
LangChain4j is an open-source Java library