i7j-rups
Apache PDFBox
i7j-rups | Apache PDFBox | |
---|---|---|
3 | 26 | |
248 | 2,395 | |
0.8% | 1.6% | |
5.3 | 9.7 | |
12 days ago | 4 days ago | |
Java | Java | |
GNU General Public License v3.0 or later | Apache License 2.0 |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
i7j-rups
-
So you want to modify the text of a PDF by hand
Great post. I've spend a lot of time reading through the PDF specification over the last ~5 years while building DocSpring [1], and I still feel like I've barely scratched the surface. qpdf is a great tool. One of my other favorites is RUPS [2], which really lets you dig into the structure of a PDF.
[1] https://docspring.com
[2] https://github.com/itext/i7j-rups
-
Show HN: I am building a new Python library to read/write PDF files
> find a version of iText RUPS application from somewhere on the internet
You mean this, right? https://github.com/itext/i7j-rups#readme
-
Any decent free online tool which can give me a breakdown of pdf contents including relative sizes of assets such as images, fonts, etc?
It's not an online tool, but it's free nonetheless: https://github.com/itext/i7j-rups
Apache PDFBox
-
PDF rendering server-side using HTML 5 + CSS 3
Are you looking for a way to render PDF's or produce them? If you want to produce PDF's, I've used https://pdfbox.apache.org/ successfully as well as https://itextpdf.com/ (potentially costs money).
-
So you want to modify the text of a PDF by hand
If you don't mind using java, you can use the open source Apache PDFBox library
https://pdfbox.apache.org/
It's relatively performant and it's a mature and supported codebase that can accomplish most pdf tasks.
- best pdf library to use in 2023?
-
How to crop, split, remove pages from PDFs with Java and PDFBox
Then, open the pdf_utils/pom.xml file and add a dependency to PDFBox, in the dependencies section:
- Does no one use PDF files anymore?? In need of a PDF generator package...
-
How to take input from User and make a PDF of it and directly send it to WhatsApp?
There are some libraries for Java that can help you create a PDF file such as PDFBox or IText. Here there's a short exaplanation on how to use them.
- Thoughts on Birt Report for pdf reports
-
How I archived 100 million PDF documents... (Part 1)
So, when I started to view the documents, a lot of them simply failed to open. I had to look around for a library that could verify PDF documents. I had some experience with PDFBox in the past, so it seemed to be a good go-to solution. It had no way to verify documents by default, but it could open and parse them and that was enough to filter out the incorrect ones. It felt a little bit strange just to read the whole PDF into the memory to verify if it is correct or not, but hey I needed a simple fix for now and it worked really well.
- Best FOSS (ideally Docker) that can split PDF files ?
-
PDF processing and analysis with open-source tools
PDFBox can do this. It’s not part of the CLI but it wouldn’t be too hard to add:
https://github.com/apache/pdfbox/blob/5b00807463279f1002e245...
What are some alternatives?
PyMuPDF - PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.
iText - [DEPRECATED] Core Java Library + PDF/A, xtra and XML Worker. Only security fixes will be added — please use iText 7
pdfsyntax - A Python library to inspect and modify the internal structure of a PDF file
OpenPDF - OpenPDF is a free Java library for creating and editing PDF files, with a LGPL and MPL open source license. OpenPDF is based on a fork of iText. We welcome contributions from other developers. Please feel free to submit pull-requests and bugreports to this GitHub repository.
djot - A light markup language
Apache FOP - Apache XML Graphics FOP
annotated-pdf-spec - Collection of useful hints for implementing a PDF library
flyingsaucer - XML/XHTML and CSS 2.1 renderer in pure Java
kaitai_struct_formats - Kaitai Struct: library of binary file formats (.ksy)
Apache POI - Mirror of Apache POI
bericht - Incremental HTML to PDF converter.
Dynamic Jasper - Dynamic Reports using Jasper Reports