Voice Recognition to Text Tool
Wan2.1: Open and Advanced Large-Scale Video Generative Model
Windows GUI Automation with Python (based on text properties)
Wan2.2: Open and Advanced Large-Scale Video Generative Model
A python parametric CAD scripting framework based on OCCT
A robust, efficient, low-latency speech-to-text library
Python module for parsing semi-structured text into python tables
AI PPT Track Terminator, the strongest PPT Skill ever
Python library and CLI tool to interface with Google Translate
Python bindings for MuPDF's rendering library.
State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX
Edit PDF files with Nano Banana
Tokenizer-Free TTS for Multilingual Speech Generation
Official inference repo for FLUX.2 models
A Powerful Native Multimodal Model for Image Generation
An AI-agent skill that turns Markdown into paste-ready WeChat article
The Ren'Py Visual Novel Engine
Instant voice cloning by MIT and MyShell. Audio foundation model
Video-based AI memory library. Store millions of text chunks in MP4
Underthesea - Vietnamese NLP Toolkit
Comprehensive Markdown plugin built for Django
Open image model at the forefront of design
Run Bonsai (1-bit) and Ternary-Bonsai language models locally
A Family of Open Sourced Music Foundation Models
Industrial-level controllable zero-shot text-to-speech system