Buzz transcribes and translates audio offline
Official inference repo for FLUX.1 models
The most powerful local music generation model
Fast stable diffusion on CPU and AI PC
Native and Compact Structured Latents for 3D Generation
Lets make video diffusion practical
Text and image to video generation: CogVideoX and CogVideo
Official code base for LeWorldModel: Stable End-to-End Joint-Embedding
tiktoken is a fast BPE tokeniser for use with OpenAI's models
The official repo of Qwen chat & pretrained large language model
Contexts Optical Compression
Project Lyra: Open Generative 3D World Models
Generate Any 3D Scene in Seconds
VGGSfM: Visual Geometry Grounded Deep Structure From Motion
Fast-stable-diffusion + DreamBooth
A Pragmatic VLA Foundation Model
Official implementation of Watermark Anything with Localized Messages
ICLR2024 Spotlight: curation/training code, metadata, distribution
GPT4V-level open-source multi-modal model based on Llama3-8B
Chinese and English multimodal conversational language model
ChatGLM-6B: An Open Bilingual Dialogue Language Model
A state-of-the-art open visual language model
Example Discord bot written in Python that uses the completions API
Fine-tuning ChatGLM-6B with PEFT
A method to increase the speed and lower the memory footprint