SaaSHub helps you find the best software and product alternatives Learn more →
Pdfcpu Alternatives
Similar projects and alternatives to pdfcpu
-
WorkOS
The modern identity platform for B2B SaaS. The APIs are flexible and easy-to-use, supporting authentication, user identity, and complex enterprise features like SSO and SCIM provisioning.
-
gotenberg
A developer-friendly API for converting numerous document formats into PDF files, and more!
-
InfluxDB
Power Real-Time Data Analytics at Scale. Get real-time insights from all types of time series data with InfluxDB. Ingest, query, and analyze billions of data points in real-time with unbounded cardinality.
-
Stirling-PDF
locally hosted web application that allows you to perform various operations on PDF files
-
go-wkhtmltopdf
Go bindings for wkhtmltopdf and high-level HTML to PDF conversion interface (by adrg)
-
maroto
A maroto way to create PDFs. Maroto is inspired in Bootstrap and uses gofpdf. Fast and simple.
-
SaaSHub
SaaSHub - Software Alternatives and Reviews. SaaSHub helps you find the best software and product alternatives
pdfcpu reviews and mentions
- Show HN: A PDF Processing CLI/API Written in Go
- Show HN
-
Making a PDF that's larger than Germany
Slightly tangential: if you are hacking on PDFs, manually or otherwise, this is an incredibly useful tool: https://pdfcpu.io/ (not the author, just a user)
-
Stirling-PDF: local web application to perform various operations on PDFs
A really nice, stand-alone command line tool is pdfcpu.
-
pdfcpu v0.6.0 out! - pdfcpu.io
Check it out => https://github.com/pdfcpu/pdfcpu/releases/tag/v0.6.0
-
Marker: Convert PDF to Markdown quickly with high accuracy
I can report that the closest I've came before is with PDFMiner (https://pypi.org/project/pdfminer/) for Python. The benefit of this one is that it retains styling information, so that italics and the like can be retained, at least with some post-processing (I think one might need to convert certain CSS-classes to actual or tags).
The other option I have started looking into is the PDFCPU library for Go. It is a bit more low-level than PDFMiner, but one gets out very well structured info, that seem it might be possible to post-process quite well, for one's particular use case and PDF layouts: https://github.com/pdfcpu/pdfcpu
I also now tried the Marker tool in the OT, and it seems to do a reasonable job. It did intermingle some columns though, at least in some tricky cases such as when there were a round shaped image in between the two columns. One note is that Marker doesn't seem to retain styling like italics though.
-
PDFcpu snippet for read text of PDF file?
Of course, the best way would be to solve it via the API without CLI. But this doesn't seem to work. https://github.com/pdfcpu/pdfcpu/issues/122
- wie splittet ihr denn PDFs - ich hab hier einige - die ich zerlegen muss in Teile
- Do you know any library to make pdf in golang?
- Pdfcpu: A Go PDF Processor
-
A note from our sponsor - SaaSHub
www.saashub.com | 19 Apr 2024
Stats
pdfcpu/pdfcpu is an open source project licensed under Apache License 2.0 which is an OSI approved license.
The primary programming language of pdfcpu is Go.