layout-parser VS py-pdf-parser

Compare layout-parser vs py-pdf-parser and see what are their differences.

Our great sponsors
  • WorkOS - The modern identity platform for B2B SaaS
  • InfluxDB - Power Real-Time Data Analytics at Scale
  • SaaSHub - Software Alternatives and Reviews
layout-parser py-pdf-parser
6 2
4,438 332
3.3% -
0.0 4.8
about 2 months ago 9 days ago
Python Python
Apache License 2.0 MIT License
The number of mentions indicates the total number of mentions that we've tracked plus the number of user suggested alternatives.
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.

layout-parser

Posts with mentions or reviews of layout-parser. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2023-01-06.

py-pdf-parser

Posts with mentions or reviews of py-pdf-parser. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2021-11-02.

What are some alternatives?

When comparing layout-parser and py-pdf-parser you can also consider the following projects:

EasyOCR - Ready-to-use OCR with 80+ supported languages and all popular writing scripts including Latin, Chinese, Arabic, Devanagari, Cyrillic and etc.

pdfplumber - Plumb a PDF for detailed information about each char, rectangle, line, et cetera — and easily extract text and tables.

tika-python - Tika-Python is a Python binding to the Apache Tika™ REST services allowing Tika to be called natively in the Python community.

BCNet - Deep Occlusion-Aware Instance Segmentation with Overlapping BiLayers [CVPR 2021]

Maya - Datetimes for Humans™

ssd_keras - A Keras port of Single Shot MultiBox Detector

mexican-government-report - Text Mining on the 2019 Mexican Government Report, covering from extracting text from a PDF file to plotting the results.

simpletransformers - Transformers for Information Retrieval, Text Classification, NER, QA, Language Modelling, Language Generation, T5, Multi-Modal, and Conversational AI

shabby-pages - ShabbyPages is a state-of-the-art corpus of born-digital document images with both ground truth and distorted versions appropriate for use in training models to reverse distortions and recover to original denoised documents.

pydantic - Data validation using Python type hints