pdf2docx

pdf2docx

Artifex
+
+

Related Products

  • Foxit Document Workflow APIs
    8 Ratings
    Visit Website
  • Adobe Acrobat
    8,334 Ratings
    Visit Website
  • MobiPDF
    7,811 Ratings
    Visit Website
  • RAD PDF
    3 Ratings
    Visit Website
  • Nutrient SDK
    113 Ratings
    Visit Website
  • Docmosis
    51 Ratings
    Visit Website
  • LM-Kit.NET
    29 Ratings
    Visit Website
  • MindCloud
    83 Ratings
    Visit Website
  • Crowdin
    944 Ratings
    Visit Website
  • Gaffa
    5 Ratings
    Visit Website

About

Create a PDF from Microsoft Office documents, protect the content, and convert to other formats. Programmatically alter a document, such as reordering, inserting, and rotating pages, as well as compressing the file. Access the same cloud-based APIs that power Adobe's end-user applications to quickly deliver scalable, secure solutions. Extract text, images, tables, and more from native and scanned PDFs into a structured JSON file. PDF Extract API leverages AI technology to accurately identify text objects and understand the natural reading order of different elements such as headings, lists, and paragraphs spanning multiple columns or pages. Extract font styles with identification of metadata such as bold and italic text and their position within your PDF. The extracted content is output in a structured JSON file format with tables in CSV or XLSX and images saved as PNG.

About

pdf2docx is a Python library that uses PyMuPDF to extract data from PDF files, parse their layouts according to rules, and generate corresponding .docx files via python-docx. It supports conversion of text, images, tables, and other structural elements; it includes tools to extract tables, handle formatting, and preserve layout as much as possible. It offers both a command-line interface and a graphical user interface. The internal architecture is modular; it includes packages for handling pages, layout, tables, images, shape paths, text spans/blocks, and other elements, enabling fine control over how PDF content is mapped into Word documents. Developers can use the API for batch conversions or integrate it into workflows; there's documentation on installation (from PyPI or source), usage, and technical details of layout-parsing, table extraction, and internal modules. The project is open source, hosted on GitHub, and made available under its license with no warranty.

Platforms Supported

Windows Not Supported
Mac Not Supported
Linux Not Supported
Cloud Supported
On-Premises Not Supported
iPhone Not Supported
iPad Not Supported
Android Not Supported
Chromebook Not Supported

Platforms Supported

Windows Not Supported
Mac Not Supported
Linux Not Supported
Cloud Supported
On-Premises Not Supported
iPhone Not Supported
iPad Not Supported
Android Not Supported
Chromebook Not Supported

Audience

Businesses searching for a PDF solution that helps create, convert, transform, OCR PDFs and more

Audience

Technical users seeking a solution to convert PDF documents into Word format programmatically while preserving layout, tables, images, and text structure

Support

Phone Support Not Supported
24/7 Live Support Not Supported
Online Supported

Support

Phone Support Supported
24/7 Live Support Not Supported
Online Supported

API

Offers API Supported

API

Offers API Supported

Screenshots and Videos

Screenshots and Videos

Pricing

No information available.
Free Version Not Supported
Free Trial Supported

Pricing

Free
Free Version Supported
Free Trial Not Supported

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation Supported
Webinars Not Supported
Live Online Not Supported
In Person Not Supported

Training

Documentation Supported
Webinars Not Supported
Live Online Not Supported
In Person Supported

Company Information

Adobe
Founded: 1982
United States
developer.adobe.com/document-services/apis/pdf-services/

Company Information

Artifex
Founded: 1993
United States
pdf2docx.readthedocs.io/en/latest/

Alternatives

Alternatives

PDF Conversa

PDF Conversa

ASCOMP Software
Pdftools

Pdftools

PDF Tools
AnyParser

AnyParser

CambioML
pdfRest API Toolkit

pdfRest API Toolkit

Datalogics Inc.
PDF.co

PDF.co

ByteScout
PDF.co

PDF.co

ByteScout

Categories

PDF APIs Supported

Categories

PDF Supported

Integrations

Microsoft Word Supported
Python Supported
.NET Supported
Adobe Acrobat Supported
Amazon Supported
EximiousSoft ePage Creator Supported
Ivo Supported
JSON Supported
Java Supported
Microsoft 365 Supported
Microsoft Excel Supported
Microsoft PowerPoint Supported
Node.js Supported
PyMuPDF Not Supported
QCommission Supported
Studiovity Screenwriting Supported
Torvalds Supported
UiPath Supported
Workers by Delos Supported

Integrations

Microsoft Word Supported
Python Supported
.NET Not Supported
Adobe Acrobat Not Supported
Amazon Not Supported
EximiousSoft ePage Creator Not Supported
Ivo Not Supported
JSON Not Supported
Java Not Supported
Microsoft 365 Not Supported
Microsoft Excel Not Supported
Microsoft PowerPoint Not Supported
Node.js Not Supported
PyMuPDF Supported
QCommission Not Supported
Studiovity Screenwriting Not Supported
Torvalds Not Supported
UiPath Not Supported
Workers by Delos Not Supported
Claim Adobe PDF Services API and update features and information
Claim Adobe PDF Services API and update features and information
Claim pdf2docx and update features and information
Claim pdf2docx and update features and information