Flexible Photo Recrafting While Preserving Your Identity
Native and Compact Structured Latents for 3D Generation
High-Resolution 3D Assets Generation with Large Scale Diffusion Models
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
Generate high-definition story short videos with one click using AI
InvokeAI is a leading creative engine for Stable Diffusion models
AV1 Image File Format Specification - ISO-BMFF/HEIF derivative
Run a full local LLM stack with one command using Docker
Open-Sora: Democratizing Efficient Video Production for All
Edit Banana: A framework for converting statistical figures
Recovering the Visual Space from Any Views
State-of-the-art Image & Video CLIP, Multimodal Large Language Models
Official MiniMax Model Context Protocol (MCP) server
Offline inference engine for art, real-time voice conversations
All-in-one native macOS AI chat application
Advanced AI Explainability for computer vision
Make any agent harness multimodal-native
State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX
Official Python inference and LoRA trainer package
Open source multimodal creative AI assistant with infinite canvas tool
We write your reusable computer vision tools
HivisionIDPhotos: a lightweight and efficient AI ID photos tools
The repository provides code for running inference with SAM 2
The electronic structure package for quantum computers
HunyuanVideo: A Systematic Framework For Large Video Generation Model