Model export recipes, Python primitives, and Swift runtime utilities
Official repository for LTX-Video
Implementation of "MobileCLIP" CVPR 2024
Python inference and LoRA trainer package for the LTX-2 audio–video
Flux 2 image generation model pure C inference
AI cognitive-enhancement Skills based on Anthropic's J-space
C#/.NET binding of llama.cpp, including LLaMa/GPT model inference
PyTorch code and models for the DINOv2 self-supervised learning
Python SDK for Claude Agent
Gemma 4 26B-A4B inference in ~2 GB of RAM on any M-series MacBook
tiktoken is a fast BPE tokeniser for use with OpenAI's models
26m function call model that runs on incredibly small devices
Open sidebar foundation, supports third-party extensions
Foundation Models for Time Series
An expressive, efficient attention architecture
Unified Multimodal Understanding and Generation Models
Multimodal embedding and reranking models built on Qwen3-VL
Instructions on how to use the Realtime API on Microcontrollers
Generate Any 3D Scene in Seconds
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
RGBD video generation model conditioned on camera input
Large-language-model & vision-language-model based on Linear Attention
Stable Diffusion with Core ML on Apple Silicon
Towards Real-World Vision-Language Understanding
Real-time behaviour synthesis with MuJoCo, using Predictive Control