apollo
ripgrep-all
apollo | ripgrep-all | |
---|---|---|
7 | 43 | |
1,361 | 6,177 | |
- | - | |
0.0 | 8.0 | |
6 months ago | 2 months ago | |
Go | Rust | |
MIT License | GNU General Public License v3.0 or later |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
apollo
- GitHub - amirgamil/apollo: A Unix-style personal search engine and web crawler for your digital footprint.
-
APSE – A Personal Search Engine
I’ve seen a handful of this kind of “Google, but only for things I’ve seen before” app. I think it’s something the world needs, but there are a lot of different approaches and I don’t think anyone has quite nailed it.
Ultimately the best solutions will likely use many different cataloging strategies depending on the content, and will allow you to tag or otherwise organize important content.
Funny enough if I had such an app I could make a list of 4 or 5 apps, but right now can only find one:
https://github.com/amirgamil/apollo
I remember seeing one posted to HN that used the browser API to essentially dump all resources on every website you visit to disk.
There’s also bookmarking and archiving tools like:
https://pinboard.in/
https://unmark.it/
https://www.linkace.org/
https://archivy.github.io/
https://github.com/kanishka-linux/reminiscence
https://perkeep.org/
- A Unix-style personal search engine and web crawler for your digital footprint
-
I built a personal search engine for my digital footprint!
Search my footprint: https://apollo.amirbolous.com/
ripgrep-all
- Ripgrep-all: rga: ripgrep, but also search PDFs, E-Books, Office documents, zip
-
Ripgrep is faster than {grep, ag, Git grep, ucg, pt, sift}
I searched in portage, and it seems there is another version working also with other documents like PDFs and doc.
https://github.com/phiresky/ripgrep-all
-
Calibre – New in Calibre 7.0
If you want even faster search across different formats, you can try ripgrep-all ( https://github.com/phiresky/ripgrep-all ). It can search across epub, docx, pdf, zip, mp4 etc. If you are handy with the tool, you can write custom adaptor to search across images using OCR with tesseract.
- Rga: Ripgrep, but also search in PDF, ebooks, office documents, zip, tar.gz etc.
-
Show HN: Khoj – Chat Offline with Your Second Brain Using Llama 2
1. If you want better adoption especially among corporations, GPL-3 wont cut it. Maybe think of some business friendly licenses (MIT etc)
2. I understand the excitement about llm's. But how about making something more accessible. I use rip-grep-all (rga) along with fzf [1] that can search all files including pdfs in a specific folders. However, I would like a GUI tool to search across multiple folders, provide priority of results across folders and store and search histories where I can do a meta-search. This is sufficient for 95% of my usecases to search locally and I dont need LLM. If khoj can enable such search as default without LLM that will be a gamechanger for many people without a heavy compute machine or who dont want to use OpenAI.
[1] https://github.com/phiresky/ripgrep-all/wiki/fzf-Integration
-
How to make file paths clickable?
I use `rga` to search through multiple PDF files for work. The tool returns a list of files and I would like to make those file paths clickable.
- Burgr – Books in Your Terminal
-
Is there a way to searching multiple epub and pdf?
rga, aka ripgrep-all
-
Internet Archive Scholar
I wanted to say 'au contrer' to your 'screenshots are not searchable' and link this[0] but I don't actually see images in the readme.. I swear it was there, maybe it's a buried extra flag..
[0] https://github.com/phiresky/ripgrep-all
- Recoll – Full-text search for your desktop
What are some alternatives?
falcon - Chrome extension for full text history search!
pdfgrep - PDFGrep is a GNU/Emacs module providing grep comparable facilities but for PDF files
dogsheep-beta - Build a search index across content from multiple SQLite database tables and run faceted searches against it using Datasette
OCRmyPDF - OCRmyPDF adds an OCR text layer to scanned PDF files, allowing them to be searched
Camlistore - Perkeep (née Camlistore) is your personal storage system for life: a way of storing, syncing, sharing, modelling and backing up content.
InvoiceNet - Deep neural network to extract intelligent information from invoice documents.
go-find-hexagonal - Applying what I learned from https://www.youtube.com/watch?v=oL6JBUk6tj0
notational-fzf-vim - Notational velocity for vim.
unmark - An open source to do app for bookmarks.
fd - A simple, fast and user-friendly alternative to 'find'
parser - 📜 Extract meaningful content from the chaos of a web page
ripgrep - ripgrep recursively searches directories for a regex pattern while respecting your gitignore