OCR software, free and offline
A Lightweight Face Recognition and Facial Attribute Analysis
Deepfakes Software For All
Autonomous research from idea to paper. Chat an Idea. Get a Paper 🦞
Wan2.1: Open and Advanced Large-Scale Video Generative Model
MiniMax H3 is a general-purpose, omni-modal generative system
AI coding assistant skill (Claude Code, Codex, OpenCode, OpenClaw)
Ready-to-use OCR with 80+ supported languages
Effortless data labeling with AI support from Segment Anything
NVR with realtime local object detection for IP cameras
A framework for the creation of autonomous agent services
Qwen3-TTS is an open-source series of TTS models
Library for OCR-related tasks powered by Deep Learning
AI Fully Automated Short Video Engine
agentUniverse is a LLM multi-agent framework
1 min voice data can also be used to train a good TTS model
Trainable models and NN optimization tools
Framework for Telegram Bot API written in Python 3.7 with asyncio
A simple, high-quality voice conversion tool focused on ease of use
Build cross-modal and multimodal applications on the cloud
Build effective agents using Model Context Protocol
LTX-Video Support for ComfyUI
A Model Context Protocol (MCP) server that enables AI assistants
A set of ready to use Agent Skills for research, science, engineering
Conversational voice AI agents