wget-lua

Wget-AT is a modern Wget with Lua hooks, Zstandard (+dictionary) WARC compression and URL-agnostic deduplication. (by ArchiveTeam)

Wget-lua Alternatives

Similar projects and alternatives to wget-lua

  1. ArchiveBox

    🗃 Open source self-hosted web archiving. Takes URLs/browser history/bookmarks/Pocket/Pinboard/etc., saves HTML, JS, PDFs, media, and more...

  2. SaaSHub

    SaaSHub - Software Alternatives and Reviews. SaaSHub helps you find the best software and product alternatives

    SaaSHub logo
  3. browsertrix-crawler

    Run a high-fidelity browser-based web archiving crawler in a single Docker container

  4. os

    11 wget-lua VS os

    Discontinued Tiny Linux distro that runs the entire OS as Docker containers

  5. libarchive

    Multi-format archive and compression library

  6. bitextor

    Bitextor generates translation memories from multilingual websites

  7. fetchurls

    A bash script to spider a site, follow links, and fetch urls (with built-in filtering) into a generated text file.

  8. grab-site

    35 wget-lua VS grab-site

    The archivist's web crawler: WARC output, dashboard for all crawls, dynamic ignore patterns

  9. 7-Zip-zstd

    7-Zip with support for Brotli, Fast-LZMA2, Lizard, LZ4, LZ5 and Zstandard

  10. pgBackRest

    Discontinued Reliable PostgreSQL Backup & Restore

  11. Crawly

    2 wget-lua VS Crawly

    Crawly, a high-level web crawling & scraping framework for Elixir.

NOTE: The number of mentions on this list indicates mentions on common posts plus user suggested alternatives. Hence, a higher number means a better wget-lua alternative or higher similarity.

wget-lua discussion

Log in or Post with

wget-lua reviews and mentions

Posts with mentions or reviews of wget-lua. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2023-02-10.
  • Alternative to HTTrack (website copier) as of 2023?
    4 projects | /r/DataHoarder | 10 Feb 2023
    You're using it wrong, rtfm, wget is still the standard. It's also extensible beyond the base feature set, take for example wget-lua ArchiveTeams well maintained go to for near all scraping projects by the group.
  • Kiwix - Access Wikipedia (And More) With no Internet
    1 project | /r/selfhosted | 4 Dec 2021
    There are updates changed names but still use more frequent updates than the dumps to get started. I know there is kiwix and xowa. Could probably build it up to current and use wget-at to scrap wikipedia solo. If you want it in html it'll proboy only be a hundred 100TB give or take. I'm wondering if any of the groups are still active on IRC. Saw mentions of a few but I lost my place in all the mobile chrome tabs.

Stats

Basic wget-lua repo stats
2
137
2.5
5 months ago

Sponsored
SaaSHub - Software Alternatives and Reviews
SaaSHub helps you find the best software and product alternatives
www.saashub.com