MiniMax H3 is a general-purpose, omni-modal generative system
AI Fully Automated Short Video Engine
Inference script for Oasis 500M
Official Python inference and LoRA trainer package
Agent Skill for generating 2D sprite sheets and map, transparent PNG
ComfyUI wrapper nodes for WanVideo and related models
A Customizable Image-to-Video Model based on HunyuanVideo
Multimodal Diffusion with Representation Alignment
Streaming Real-time Audio-Driven Avatar Generation
AI logo animation skill: turn raster logos into smooth SVG animation
Positron, a next-generation data science IDE
Qwen2.5-VL is the multimodal large language model series
OCR expert VLM powered by Hunyuan's native multimodal architecture
A speech-text foundation model for real time dialogue
A Customizable Image-to-Video Model based on HunyuanVideo
Image & video editor with clarity, sharpness & fast FFmpeg export