P
PyMuPDF
pymupdf/PyMuPDF
PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.
★10.4kstars
Python
AGPL-3.0
Updated: Today
📋 Project at a Glance
Tap to expand
What's this?A , built with Python open-source project in the Data & Infrastructure category, core strengths: data-science/epub
Who made it?Maintained by pymupdf team, 10.4K⭐ on GitHub, #187 out of 3133 in Data & Infrastructure
Why does it exist?As the Data & Infrastructure landscape evolved, the pymupdf team identified the need for better data-science solutions. PyMuPDF was created to simplify epub workflows.
What can it do?Key use cases: extract-data, font, mupdf
How to install with AI?Use an AI coding assistant (Claude Code, Cursor, Copilot) to automatically set up pip dependencies and virtual env. Follow the README — the AI handles the rest.
🔗 github.com/pymupdf/PyMuPDF | 官网 https://pymupdf.readthedocs.io/?utm_source=github&utm_medium=referral&utm_campaign=pymupdf_github&utm_content=about&utm_term=docs
🔗 github.com/pymupdf/PyMuPDF | 官网 https://pymupdf.readthedocs.io/?utm_source=github&utm_medium=referral&utm_campaign=pymupdf_github&utm_content=about&utm_term=docs
Topics
data-scienceepubextract-datafontmupdfocrpdfpdf-documentspymupdfpythontable-extractiontesseracttext-processingtext-shapingxps