Model export recipes, Python primitives, and Swift runtime utilities
Official repository for LTX-Video
Native and Compact Structured Latents for 3D Generation
Implementation of "MobileCLIP" CVPR 2024
C#/.NET binding of llama.cpp, including LLaMa/GPT model inference
Python inference and LoRA trainer package for the LTX-2 audio–video
Flux 2 image generation model pure C inference
AI cognitive-enhancement Skills based on Anthropic's J-space
PyTorch code and models for the DINOv2 self-supervised learning
Python SDK for Claude Agent
Gemma 4 26B-A4B inference in ~2 GB of RAM on any M-series MacBook
tiktoken is a fast BPE tokeniser for use with OpenAI's models
26m function call model that runs on incredibly small devices
Foundation Models for Time Series
An expressive, efficient attention architecture
Open sidebar foundation, supports third-party extensions
Unified Multimodal Understanding and Generation Models
State-of-the-art Image & Video CLIP, Multimodal Large Language Models
Multimodal embedding and reranking models built on Qwen3-VL
Instructions on how to use the Realtime API on Microcontrollers
Generate Any 3D Scene in Seconds
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
RGBD video generation model conditioned on camera input
Large-language-model & vision-language-model based on Linear Attention
Towards Real-World Vision-Language Understanding