MiniMax H3 inference engine for Mac computers
TT-NN operator library, and TT-Metalium low level kernel programming
Metal programming in Julia
A bare metal raycaster, boots from a floppy image
Heavy metal SOAP client
Run serverless GPU workloads with fast cold starts on bare-metal
DeepSeek 4 Flash local inference engine for Metal
General purpose Linux OS for Azure
A network load-balancer implementation for Kubernetes
Zero Hour running natively on macOS, iPhone & iPad
Complete BIOS and firmware packs for RetroArch, Batocera, Recalbox
Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference
Gemma 4 26B-A4B inference in ~2 GB of RAM on any M-series MacBook
Use Autodesk Fusion 360 on Linux
Deploy a Production Ready Kubernetes Cluster
Productive, portable, and performant GPU programming in Python
Fastest LLM inference runtime for Apple Silicon
A scalable inference server for models optimized with OpenVINO
Python implementation for microcontrollers and constrained systems
A modern cross-platform low-level graphics API
A Rust-based, lightweight unikernel
Safe and portable GPU abstraction in Rust, implementing WebGPU API
High-performance TensorFlow Lite library for React Native
A template for deploying a Kubernetes cluster with k3s or Talos
Open source hyperconverged infrastructure (HCI) software