Open source personal AI Assistant for Linux, Windows and Mac
Foundational video generation model with 13.6B parameters
21 Lessons, Get Started Building with Generative AI
Stable Diffusion built-in to Blender
Generate Any 3D Scene in Seconds
Pretrained model hub for Keras 3
Official Python inference and LoRA trainer package
Sample code and notebooks for Generative AI on Google Cloud
Phi-3.5 for Mac: Locally-run Vision and Language Models
Open-Sora: Democratizing Efficient Video Production for All
The data structure for multimodal data
Large-language-model & vision-language-model based on Linear Attention
HunyuanVideo: A Systematic Framework For Large Video Generation Model
Official implementation of DreamCraft3D
Pushing the Frontier of Long Audio-Visual Generation
Make any agent harness multimodal-native
A Systematic Framework for Interactive World Modeling
litellm without the bloat
Open source libraries and APIs to build custom preprocessing pipelines
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
Framework for building neural networks
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
Infinite Worlds with Versatile Interactions
Gemma open-weight LLM library, from Google DeepMind
Integrate ChatGPT into your own discord bot