reddit-save
ArchiveBox
Our great sponsors
reddit-save | ArchiveBox | |
---|---|---|
6 | 248 | |
125 | 19,737 | |
- | 3.1% | |
4.4 | 9.7 | |
4 months ago | 12 days ago | |
Python | Python | |
- | MIT |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
reddit-save
-
Download Saved Reddit posts into Obsidian automatically in .md format
This was the base of it: https://github.com/samirelanduk/reddit-save
- What is a self-hostable service you wish existed?
-
Is there a good list of up-to-date data archiving tools for different websites?
I'm mostly on reddit and I use reddit-save. It works well! Biggest issue is I'd like to be able to archive a thread to an arbitrary length.
- I fear this sub will be quarantined soon. Is there a way to archive stuff?
-
Limiting Access to Removed and Deleted Post Pages
reddit-save
-
Starting my own Data Hoarding
Reddit: yes, we are on reddit and there is considerable amount of science and art. Already archiving all saved posts using reddit-save and some low-traffic subreddits in their entirety using unknown tool.
ArchiveBox
-
Ask HN: What Underrated Open Source Project Deserves More Recognition?
Two projects I greatly appreciate, allowing me to easily archive my bandcamp and GOG purchases (after the initial setup anyways):
https://github.com/easlice/bandcamp-downloader
https://github.com/Kalanyr/gogrepoc
And I recently learned about archivebox, which I think is going to be a fast favorite and finally let me clear out my mess of tabs/bookmarks: https://github.com/ArchiveBox/ArchiveBox
- YaCy, a distributed Web Search Engine, based on a peer-to-peer network
-
Vice website is shutting down
If you really want to save the content for yourself, use something like https://archivebox.io/
I've been running a local instance for a few years now and download/save tech articles all time. I can search and find them as needed.
-
An Introduction to the WARC File
API is coming soon (relatively, it's still a one-man project)! Stay tuned https://github.com/ArchiveBox/ArchiveBox/issues/496
I have an event-sourcing refactor in progress now to allow us to pluginize functionality like the API (similar to Home Assistant with a plugin app sotre), it will take a month or two. Next up is the REST API using the new plugin system.
-
Ask HN: How can I back up an old vBulletin forum without admin access?
I guess your best chance is to use something like https://archivebox.io/.
-
ArchiveBox – open-source self-hosted web archiving
Yeah this is a cool project but it was discussed 2 days ago.
As mentioned by the maintainer there, they even maintain a list of alternatives, very classy:
https://github.com/ArchiveBox/ArchiveBox/wiki/Web-Archiving-...
- ArchiveBox: Open-source self-hosted web archiving
- Linkhut: A Social Bookmarking Site
- Show HN: Rem: Remember Everything (open source)
- Bookmark manager with a focus on organization?
What are some alternatives?
bulk-downloader-for-reddit - Downloads and archives content from reddit
Wallabag - wallabag is a self hostable application for saving web pages: Save and classify articles. Read them later. Freely.
TumblThree - A Tumblr and Twitter Blog Backup Application
paimon-moe - Your best Genshin Impact companion! Help you plan what to farm with ascension calculator and database. Also track your progress with todo and wish counter.
youtube-dl - Command-line program to download videos from YouTube.com and other video sites
SingleFile - Web Extension for saving a faithful copy of a complete web page in a single HTML file
wikiteam - Tools for downloading and preserving wikis. We archive wikis, from Wikipedia to tiniest wikis. As of 2023, WikiTeam has preserved more than 350,000 wikis.
ArchivesSpace - The ArchivesSpace archives management tool
reveddit - Review removed content on reddit. Uses the Pushshift API, built on code from removeddit.
grab-site - The archivist's web crawler: WARC output, dashboard for all crawls, dynamic ignore patterns
floccus - :cloud: Sync your bookmarks privately across browsers and devices
Archivematica - Free and open-source digital preservation system designed to maintain standards-based, long-term access to collections of digital objects.