Enhances Tesseract OCR output using LLMs (local or API)
Workflow and speech recognition app
A Repo For Document AI
General-purpose image editing model that delivers high-fidelity
Open speech-to-speech models and pipelines by Hugging Face toolkit AI
Qwen3-ASR is an open-source series of ASR models
Google Gen AI Python SDK provides an interface for developers
Stable Diffusion web UI
Private Open AI on Kubernetes
LLM-based agent for general purpose software engineering tasks
Agent Skill for generating 2D sprite sheets and map, transparent PNG
Semantic Search & Call Graphs for AI Agents
A python tool that uses GPT-4, FFmpeg, and OpenCV
Running large language models on a single GPU
ChatGPT extension for scientific research work
Create, Edit, Delete, Organize , Convert, Export, Secure & Sign PDF.
Translate English to Bangla using CSV file format and range wise.
Free, open-source, offline speech-to-text tool for Windows and MacOS.
Download books from the hathitrust website in a fast and easy manner
AI-powered semantic indexing: automating the creation of book indexes
Create software using a general-purpose visual programming system
Free, open-source Opus Clip alternative for Windows.
Download, transcribe and convert videos on your PC. No cloud.
One-click deployment (including offline integration package)