PyMuPDF

by pymupdf · Data

PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other)

New0 ratings10,659 starsActive
DataDeveloper ToolPythonSWIG
Open project ↗↓ Download 1.28.2GitHub⚑ Report
PyMuPDF preview

About this project

PyMuPDF The PDF engine behind over 50 million monthly downloads, powering AI pipelines worldwide. PyMuPDF is a high-performance Python library for data extraction, analysis, conversion, rendering and manipulation of PDF (and other) documents. Built on top of MuPDF — a lightweight, fast C engine — PyMuPDF gives you precise, low-level control over documents alongside high-level convenience APIs. No mandatory external dependencies. Why PyMuPDF? — Fast — powered by MuPDF, a best-in-class C rendering engine — Accurate — pixel-perfect text extraction with font, color, and position metadata — Versatile — read, write, annotate, redact, merge, split, and…

Technologies

PythonHTMLCdata-scienceSWIGepubextract-data

Project health

Actively maintained
Last updateyesterday
Contributors92
Latest release1.28.2
Open issues & PRs57
LicenseAGPL-3.0
On GitHubsince 2012

GitHub

10,659
stars
796
forks
92
contributors
57
open issues & PRs
Python
language
yesterday
last commit
View on GitHub ↗All releases ↗

Reviews

out of 5 · 0 ratings
★★★★★
0%
★★★★
0%
★★★
0%
★★
0%
0%
Sign in to write a review
No reviews yet
Be the first to review PyMuPDF.

Built by

pymupdf
Imported from GitHub · not yet claimed on GitPalace
View developer pageSign in with GitHub to claim

Maintain pymupdf/PyMuPDF? Claiming verifies admin access through your GitHub account and gives you control of this listing.

You might also like