PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.
-
Updated
Aug 21, 2026 - Python
PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.
Shape-aware Unicode normalizer for Traditional Mongolian (Hudum): same visible word ⟹ same encoding. Pure-Python UTN #57 v4 shaping engine, zero dependencies. ᠮᠣᠩᠭᠣᠯ ᠪᠢᠴᠢᠭ
PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.
A comprehensive Python tool for analyzing PDF files and determining the best PDF processing library for each file. The analyzer tests PDFs against multiple libraries (pypdf, PyMuPDF, pdfplumber) and provides detailed compatibility reports and recommendations.
Add a description, image, and links to the text-shaping topic page so that developers can more easily learn about it.
To associate your repository with the text-shaping topic, visit your repo's landing page and select "manage topics."