Buzz transcribes and translates audio offline
Official inference repo for FLUX.1 models
Flux 2 image generation model pure C inference
Port of Facebook's LLaMA model in C/C++
super expressive prompting model based on ltx2.3
The most powerful local music generation model
Run the full 2.78-trillion-parameter Kimi K3 model
Fast stable diffusion on CPU and AI PC
Lets make video diffusion practical
Instructions on how to use the Realtime API on Microcontrollers
Text and image to video generation: CogVideoX and CogVideo
Official code base for LeWorldModel: Stable End-to-End Joint-Embedding
tiktoken is a fast BPE tokeniser for use with OpenAI's models
Contexts Optical Compression
Project Lyra: Open Generative 3D World Models
Generate Any 3D Scene in Seconds
Fast-stable-diffusion + DreamBooth
Continuous Autonomy for the AI SDK
ICLR2024 Spotlight: curation/training code, metadata, distribution
Blazeface is a lightweight model that detects faces in images
ChatGLM-6B: An Open Bilingual Dialogue Language Model
A CNN model that predicts human joints from RGB images of a person
Detect faces in an image
A state-of-the-art open visual language model
Python example app from the OpenAI API quickstart tutorial