AO3_Scraper
dude
AO3_Scraper | dude | |
---|---|---|
2 | 28 | |
7 | 412 | |
- | - | |
7.6 | 9.0 | |
6 months ago | 6 days ago | |
Python | Python | |
MIT License | GNU Affero General Public License v3.0 |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
AO3_Scraper
- And I thought amazing fics suddenly being deleted was a myth
-
scrape bookmarks for spreadsheet
Do you still need help with this? The program Calibre along with the plugin FanFicFare can do that (this requires a bit of setup, but is useful because it downloads the fics and makes the tag info easily filterable in the Calibre GUI). Or you can use something like AO3_Scraper if you just want a spreadsheet.
dude
-
Webscraping beginner here ready to start leveling up to intermediate. Looking for some good webscraping repositories (e.g any of your GitHub repos/projects) that I can use as learning tools, and general recommendations for what to do next
Please check https://github.com/roniemartinez/dude
-
Need help with downloading a section of multiple sites as pdf files.
You can use my library which also uses Playwright. I have an example here: https://github.com/roniemartinez/dude/discussions/116
-
Why do you use python for web scraping?
I also built a framework so I can easily switch between these libraries with less code change (still on hiatus for a few months before going back to it): https://github.com/roniemartinez/dude
-
Thank GOD for Poetry!
There's a lot of options but I am quite happy with Github Actions workflows + Poetry as it handles tests and publish to PyPI. Just an example, in my workflows, I deploy to TestPyPI and PyPI here: https://github.com/roniemartinez/dude/tree/master/.github/workflows
-
What stack or tools are you using for ensuring code quality and best practices in medium and large codebases ?
But for documentation, I use mkdocs-material as it can easily be used with minor customization and changes can be easily deployed in Github: https://roniemartinez.github.io/dude/
- Is there any thing Beautifulsoup can do that Scrapy can not?
-
Screenshotting site, but remove all popups.
Add an adblocker. I implemented Dude/pydude with the this and page results are clean without ads and pop-ups. For the screenshot, here is an example: https://github.com/roniemartinez/dude/discussions/116
-
which Python Library is best for scraping?
You can also use my library if you want things to be simpler:) https://github.com/roniemartinez/dude
-
For those of you using Python, what is your go to library to build your scraper?
I use my own library, Dude! https://github.com/roniemartinez/dude
-
Building a (relatively) easily adaptable, flexible web scraper (seeking conceptual advice)
I built a simple web scraper that is simple to use but this is still a work-in-progress - https://github.com/roniemartinez/dude
What are some alternatives?
outlook-account-generator - Outlook Account Generator helps you create outlook accounts.
Edu-Mail-Generator - Generate Free Edu Mail(s) within minutes
GoodreadsScraper - Scrape data from Goodreads using Scrapy and Selenium :books:
python-web-scraping-primjeri - web scraping stranica posta.hr, konzum.hr, index.hr, njuskalo.hr, neostar.com, DasWeltAuto.hr, ...
facebook_page_scraper - Scrapes facebook's pages front end with no limitations & provides a feature to turn data into structured JSON or CSV
scrapy-playwright - 🎠Playwright integration for Scrapy
autoscraper - A Smart, Automatic, Fast and Lightweight Web Scraper for Python
FastDepends - FastDepends - FastAPI Dependency Injection system extracted from FastAPI and cleared of all HTTP logic. Async and sync modes are both supported.
ffnToAO3 - Transferring your works from FFN to AO3
HomeHarvest - Python package for real estate scraping of MLS listing data [Moved to: https://github.com/Bunsly/HomeHarvest]
dnd-roll-parser - Python project that will take the saved html chat log and calculate the average rolls per player.
Cascadia.jl - A CSS Selector library in Julia