Buzz transcribes and translates audio offline
Official inference repo for FLUX.1 models
Flux 2 image generation model pure C inference
Port of Facebook's LLaMA model in C/C++
super expressive prompting model based on ltx2.3
Run the full 2.78-trillion-parameter Kimi K3 model
The most powerful local music generation model
Fast stable diffusion on CPU and AI PC
Native and Compact Structured Latents for 3D Generation
Instructions on how to use the Realtime API on Microcontrollers
Lets make video diffusion practical
Text and image to video generation: CogVideoX and CogVideo
Official code base for LeWorldModel: Stable End-to-End Joint-Embedding
tiktoken is a fast BPE tokeniser for use with OpenAI's models
The official repo of Qwen chat & pretrained large language model
Contexts Optical Compression
Project Lyra: Open Generative 3D World Models
Generate Any 3D Scene in Seconds
VGGSfM: Visual Geometry Grounded Deep Structure From Motion
Fast-stable-diffusion + DreamBooth
A Pragmatic VLA Foundation Model
Official implementation of Watermark Anything with Localized Messages
Continuous Autonomy for the AI SDK
ICLR2024 Spotlight: curation/training code, metadata, distribution
GPT4V-level open-source multi-modal model based on Llama3-8B