Page 2 | python text parser free download

pywinauto

Windows GUI Automation with Python (based on text properties)

pywinauto is a set of Python modules to automate the Microsoft Windows GUI. At its simplest it allows you to send mouse and keyboard actions to Windows dialogs and controls, but it has support for more complex actions like getting text data.

Downloads: 6 This Week

Last Update: 2025-01-06

See Project

Jupyter Notebook Tools for Sphinx

Sphinx source parser for Jupyter notebooks

nbsphinx is a Sphinx extension that provides a source parser for *.ipynb files. Custom Sphinx directives are used to show Jupyter Notebook code cells (and of course their results) in both HTML and LaTeX output. Un-evaluated notebooks – i.e. notebooks without stored output cells – will be automatically executed during the Sphinx build process.

Downloads: 1 This Week

Last Update: 4 days ago

See Project

MeloTTS

High-quality multi-lingual text-to-speech library by MyShell.ai

MeloTTS is an open-source text-to-speech (TTS) system that generates natural-sounding speech from text input. It utilizes advanced machine-learning models to produce high-quality audio outputs.

Downloads: 6 This Week

Last Update: 2025-01-06

See Project

MegaParse

File Parser optimised for LLM Ingestion with no loss

MegaParse is a file parser optimized for Large Language Model (LLM) ingestion, ensuring no loss of information. It efficiently parses various document formats, such as PDFs, DOCX, and PPTX, converting them into formats ideal for processing by LLMs. This tool is essential for applications that require accurate and comprehensive data extraction from diverse document types.

Downloads: 3 This Week

Last Update: 2025-02-14

See Project

OCRmyPDF

OCRmyPDF adds an OCR text layer to scanned PDF files

OCRmyPDF adds an optical character recognition (OCR) text layer to scanned PDF files, allowing them to be searched. PDF is the best format for storing and exchanging scanned documents. Unfortunately, PDFs can be difficult to modify. OCRmyPDF makes it easy to apply image processing and OCR (recognized, searchable text) to existing PDFs.

Downloads: 91 This Week

Last Update: 2025-11-11

See Project

JC

CLI tool and python library

...The JC parsers can also be used as python modules. In this case, the output will be a python dictionary, or a list of dictionaries, instead of JSON. Two representations of the data are available. The default representation uses a strict schema per parser and converts known numbers to int/float JSON values. Certain known values of None are converted to JSON null, known boolean values are converted, and, in some cases, additional semantic context fields are added.

Downloads: 2 This Week

Last Update: 2025-10-13

See Project

novelWriter

Open source plain text editor designed for writing novels

A markdown-like text editor designed for writing novels and larger projects of many smaller plain text documents. It is designed to be a simple text editor that allows for easy organization of text files and notes, with a metadata syntax for comments, synopsis, and cross-referencing between files, and built on plain text files for robustness. The project storage is suitable for version control software, and also well suited for file synchronisation tools. All text is saved as plain text...

Downloads: 2 This Week

Last Update: 2025-09-14

See Project

markdown-rs

CommonMark compliant markdown parser in Rust with ASTs and extensions

markdown-rs is an open-source markdown parser written in Rust. It’s implemented as a state machine (#![no_std] + alloc) that emits concrete tokens, so that every byte is accounted for, with positional info. The API then exposes this information as an AST, which is easier to work with, or it compiles directly to HTML. While most markdown parsers work towards compliancy with CommonMark (or GFM), this project goes further by following how the reference parsers (cmark, cmark-gfm) work, which is...

Downloads: 0 This Week

Last Update: 2025-04-23

See Project

SAM 3

Code for running inference and finetuning with SAM 3 model

SAM 3 (Segment Anything Model 3) is a unified foundation model for promptable segmentation in both images and videos, capable of detecting, segmenting, and tracking objects. It accepts both text prompts (open-vocabulary concepts like “red car” or “goalkeeper in white”) and visual prompts (points, boxes, masks) and returns high-quality masks, boxes, and scores for the requested concepts. Compared with SAM 2, SAM 3 introduces the ability to exhaustively segment all instances of an...

Downloads: 80 This Week

Last Update: 2025-11-25

See Project

RealtimeSTT

A robust, efficient, low-latency speech-to-text library

RealtimeSTT is a Python-based realtime speech-to-text engine emphasizing low latency, wake-word detection, voice activity detection, and automatic speech segmentation. It provides asynchronous callbacks, nanosecond-precision timestamps, and CLI tools, suitable for building voice assistants, meeting transcribers, or live caption systems.

Downloads: 3 This Week

Last Update: 2025-07-03

See Project

PaddleOCR

Awesome multilingual OCR toolkits based on PaddlePaddle

PaddleOCR offers exceptional, multilingual, and practical Optical Character Recognition (OCR) tools that can help users train better models and apply them into practice. Inspired by PaddlePaddle, PaddleOCR is an ultra lightweight OCR system, with multilingual recognition, digit recognition, vertical text recognition, as well as long text recognition. It features a PPOCR series of high-quality pre-trained models, which includes: ultra lightweight ppocr_mobile series models, general...

Downloads: 48 This Week

Last Update: 2025-11-13

See Project

PyMdown Extensions

Extensions for Python Markdown

PyMdown Extensions is a collection of extensions for Python Markdown. They were originally written to make writing documentation more enjoyable. They cover a wide range of solutions, and while not every extension is needed by all people, there is usually at least one useful extension for everybody. All extensions are found under the module namespace of pymdownx. Assuming we wanted to specify the use of the MagicLink extension, we would include it in Python Markdown.

Downloads: 1 This Week

Last Update: 6 days ago

See Project

rich

Rich is a Python library for rich text and beautiful formatting

The Rich API makes it easy to add color and style to terminal output. Rich can also render pretty tables, progress bars, markdown, syntax highlighted source code, tracebacks, and more, out of the box. Rich is a Python library for rich text and beautiful formatting in the terminal. Rich works with Linux, OSX, and Windows. True color/emoji works with new Windows Terminal, classic terminal is limited to 16 colors. Rich requires Python 3.7 or later. Effortlessly add rich output to your application, you can import the rich print method, which has the same signature as the builtin Python function. ...

Downloads: 5 This Week

Last Update: 2025-10-09

See Project

sherpa-onnx

Speech-to-text, text-to-speech, and speaker recognition

Speech-to-text, text-to-speech, and speaker recognition using next-gen Kaldi with onnxruntime without an Internet connection. Support embedded systems, Android, iOS, Raspberry Pi, RISC-V, x86_64 servers, websocket server/client, C/C++, Python, Kotlin, C#, Go, NodeJS, Java, Swift, Dart, JavaScript, Flutter.

Downloads: 23 This Week

Last Update: 1 day ago

See Project

Voice-Pro

Comprehensive Gradio WebUI for audio processing

Speech recognition module for Python

Library for performing speech recognition, with support for several engines and APIs, online and offline. Recognize speech input from the microphone, transcribe an audio file, save audio data to an audio file. Show extended recognition results, calibrate the recognizer energy threshold for ambient noise levels (see recognizer_instance.energy_threshold for details). Listening to a microphone in the background, various other useful recognizer features. The easiest way to install this is using...

Downloads: 10 This Week

Last Update: 2025-11-19

See Project

Search Results for "python text parser" - Page 2

Showing 1639 open source projects for "python text parser"

pywinauto

Jupyter Notebook Tools for Sphinx

MeloTTS

MegaParse

OCRmyPDF

JC

novelWriter

markdown-rs

SAM 3

RealtimeSTT

PaddleOCR

PyMdown Extensions

rich

sherpa-onnx

Voice-Pro

Fooocus

gTTS

Whisper

pyttsx3

PyTextRank

MLX-Audio

edge-tts

Jupytext

MarkItDown

SpeechRecognition

Search Results for "python text parser" - Page 2

Showing 1639 open source projects for "python text parser"

pywinauto

Jupyter Notebook Tools for Sphinx

MeloTTS

MegaParse

OCRmyPDF

JC

novelWriter

markdown-rs

SAM 3

RealtimeSTT

PaddleOCR

PyMdown Extensions

rich

sherpa-onnx

Voice-Pro

Fooocus

gTTS

Whisper

pyttsx3

PyTextRank

MLX-Audio

edge-tts

Jupytext

MarkItDown

SpeechRecognition

Related Searches

Related Categories