PyMuPDF vs PDFMiner

Get real-time insights from all types of time series data with InfluxDB. Ingest, query, and analyze billions of data points in real-time with unbounded cardinality.

www.influxdata.com

featured

SaaSHub - Software Alternatives and Reviews

SaaSHub helps you find the best software and product alternatives

www.saashub.com

featured

PyMuPDF		PDFMiner
	Project
5	Mentions	6
4,103	Stars	5,179
5.3%	Growth	-
9.8	Activity	0.0
3 days ago	Latest Commit	over 1 year ago
Python	Language	Python
GNU Affero General Public License v3.0	License	MIT License

The number of mentions indicates the total number of mentions that we've tracked plus the number of user suggested alternatives.
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.

PyMuPDF

Posts with mentions or reviews of PyMuPDF. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2023-12-04.

FLaNK Stack for 04 December 2023
24 projects | dev.to | 4 Dec 2023
Converting markdown to pdf in Python
3 projects | dev.to | 12 Oct 2023

This method is based on the use of the libraries markdown-it-py (conversion from markdown to html) and [PyMuPDF] https://github.com/pymupdf/PyMuPDF) (conversion from html to pdf). A small Python class links them together.
Show HN: I am building a new Python library to read/write PDF files
17 projects | news.ycombinator.com | 17 Nov 2022

I think you might mean PyMuPDF (https://github.com/pymupdf/PyMuPDF), a Python library built on top of the MuPDF C library (https://mupdf.com/).
PyMuPDF and MuPDF are both available under dual open source AGPL and commercial licenses. They have been around for many years and are under continual development.
[Disclaimer, i work for Artifex, who wrote MuPDF and recently acquired PyMuPDF.]
M1 Mac: myuPDF install (wheel?)
1 project | /r/learnpython | 4 Sep 2022
legacy install error: PyMuPDF?
1 project | /r/learnpython | 4 Sep 2022

PDFMiner

Posts with mentions or reviews of PDFMiner. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2022-10-15.

Creating a python class for organizing courses I took in my education
2 projects | /r/learnpython | 15 Oct 2022

Technically this information is on my transcript, so I will be trying to use pdfminer to extract that data if there is a way to use a class you recommend when using that code https://github.com/pdfminer/pdfminer.six
Dúvida séria sobre Metadados!
1 project | /r/brdev | 27 Jul 2022
Desktop File Search, How does it work! python or C# library libraries?
1 project | /r/learnpython | 12 Sep 2021
Add Texts to existing PDF using Python
1 project | /r/learnpython | 23 Jun 2021

PDFMiner - for getting the fields of every possible PDF : only needed for weird cases where standard PDF reader methods do not work
I'm having trouble with the PDFminer library in Python. Whenever I try to call a certain function it says that something is missing in the library itself
2 projects | /r/AskProgramming | 22 May 2021

The Github page says it's been superseded by pdfminer.six. Perhaps try that instead.
Extract specific data from multiple PDF files
1 project | /r/learnpython | 5 Mar 2021

What are some alternatives?

When comparing PyMuPDF and PDFMiner you can also consider the following projects:

PyPDF2 - A pure-python PDF library capable of splitting, merging, cropping, and transforming the pages of PDF files

ReportLab

pdfplumber - Plumb a PDF for detailed information about each char, rectangle, line, et cetera — and easily extract text and tables.

pdfminer.six - Community maintained fork of pdfminer - we fathom PDF

borb - borb is a library for reading, creating and manipulating PDF files in python.

Camelot - A Python library to extract tabular data from PDFs

pdfquery - A fast and friendly PDF scraping library.

WeasyPrint - The awesome document factory

PyMuPDF vs PyPDF2 PDFMiner vs PyPDF2 PyMuPDF vs ReportLab PDFMiner vs pdfplumber PyMuPDF vs pdfplumber PDFMiner vs pdfminer.six PyMuPDF vs borb PDFMiner vs Camelot PyMuPDF vs pdfquery PDFMiner vs WeasyPrint PyMuPDF vs WeasyPrint PDFMiner vs ReportLab

Compare PyMuPDF vs PDFMiner and see what are their differences.

PyMuPDF

PDFMiner

PyMuPDF

PDFMiner

What are some alternatives?