TT-NN operator library, and TT-Metalium low level kernel programming
MiniMax H3 inference engine for Mac computers
Metal programming in Julia
Heavy metal SOAP client
DeepSeek 4 Flash local inference engine for Metal
Zero Hour running natively on macOS, iPhone & iPad
Gemma 4 26B-A4B inference in ~2 GB of RAM on any M-series MacBook
Run x86-64 Windows PC games on jailed iOS via FEX-Emu + Wine + DXMT
Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference
General purpose Linux OS for Azure
Run serverless GPU workloads with fast cold starts on bare-metal
A network load-balancer implementation for Kubernetes
Complete BIOS and firmware packs for RetroArch, Batocera, Recalbox
Kubernetes-native system managing the full lifecycle of Kubernetes
Deploy a Production Ready Kubernetes Cluster
Use Autodesk Fusion 360 on Linux
Run frontier MoE models on hardware you already own
Productive, portable, and performant GPU programming in Python
Python implementation for microcontrollers and constrained systems
Safe and portable GPU abstraction in Rust, implementing WebGPU API
QVAC Fabric: cross-platform LLM inference and fine-tuning
Domain-specific language designed to streamline the development
High-performance TensorFlow Lite library for React Native
A scalable inference server for models optimized with OpenVINO
A Rust-based, lightweight unikernel