Phi-3.5 for Mac: Locally-run Vision and Language Models
Gemma 4 26B-A4B inference in ~2 GB of RAM on any M-series MacBook
MiniMax H3 inference engine for Mac computers
Model export recipes, Python primitives, and Swift runtime utilities
Sharp Monocular Metric Depth in Less Than a Second
Lightweight MoE model for local reasoning, coding, and AI agents
Fast uncensored Gemma model optimized for local chat and coding
Open agentic coding model optimized for local deployment