Internet-Places-Database Alternatives

Similar projects and alternatives to Internet-Places-Database

ArchiveBox

248 19,737 9.7 Python Internet-Places-Database VS ArchiveBox

🗃 Open source self-hosted web archiving. Takes URLs/browser history/bookmarks/Pocket/Pinboard/etc., saves HTML, JS, PDFs, media, and more...
Yacy

115 3,253 8.7 Java Internet-Places-Database VS Yacy

Distributed Peer-to-Peer Web Search Engine and Intranet Search Appliance
WorkOS

workos.com sponsored

The modern identity platform for B2B SaaS. The APIs are flexible and easy-to-use, supporting authentication, user identity, and complex enterprise features like SSO and SCIM provisioning.
Filestash

108 9,414 9.3 JavaScript Internet-Places-Database VS Filestash

🦄 A modern web client for SFTP, S3, FTP, WebDAV, Git, Minio, LDAP, CalDAV, CardDAV, Mysql, Backblaze, ...
Miniflux

87 6,228 9.7 Go Internet-Places-Database VS Miniflux

Minimalist and opinionated feed reader
LinkAce

48 2,426 7.7 PHP Internet-Places-Database VS LinkAce

LinkAce is a self-hosted archive to collect links of your favorite websites.
chatgpt-shell

25 764 9.3 Emacs Lisp Internet-Places-Database VS chatgpt-shell

ChatGPT and DALL-E Emacs shells + Org babel 🦄 + a shell maker for other providers
RSS-Link-Database

9 8 9.5 Internet-Places-Database VS RSS-Link-Database

Bookmarked archived links
InfluxDB

www.influxdata.com sponsored

Power Real-Time Data Analytics at Scale. Get real-time insights from all types of time series data with InfluxDB. Ingest, query, and analyze billions of data points in real-time with unbounded cardinality.
motion

16 3,545 6.3 C Internet-Places-Database VS motion

Motion, a software motion detector. Home page: https://motion-project.github.io/ (by Motion-Project)
Django-link-archive

12 11 9.6 Python Internet-Places-Database VS Django-link-archive

Link archive for a NAS drive
polychrome.nvim

1 11 6.1 Lua Internet-Places-Database VS polychrome.nvim

A colorscheme creation micro-framework for Neovim
webring

3 807 7.5 HTML Internet-Places-Database VS webring

Make yourself a website
notifeed

1 1 10.0 Python Internet-Places-Database VS notifeed

Watch RSS/Atom feeds and send push notifications/webhooks when new content is detected
soundfingerprinting

6 909 8.1 C# Internet-Places-Database VS soundfingerprinting

Open source audio fingerprinting in .NET. An efficient algorithm for acoustic fingerprinting written purely in C#.
klog

6 515 7.6 Go Internet-Places-Database VS klog

Command line tool for time tracking in a human-readable, plain-text file format. (by jotaen)
RSS-Link-Database-2023

6 2 9.4 HTML Internet-Places-Database VS RSS-Link-Database-2023

link archive for year 2023
webpub

1 9 10.0 TypeScript Internet-Places-Database VS webpub

Give me a website, I'll make you an epub.
srgn

5 385 9.5 Rust Internet-Places-Database VS srgn

A code surgeon for precise text and code transplantation. A marriage of `tr`/`sed`, `rg` and `tree-sitter`.
recess

5 16 8.1 TypeScript Internet-Places-Database VS recess

A content aggregator for keeping up and interacting with siloed content. (by yakkomajuri)
kindle_clippings_webapp

5 7 7.2 JavaScript Internet-Places-Database VS kindle_clippings_webapp

Web Application for importing, viewing and tagging kindle clippings. Account is not required.
full-text-tabs-forever

4 52 8.5 TypeScript Internet-Places-Database VS full-text-tabs-forever

Full text search all your browsing history
SaaSHub

www.saashub.com sponsored

SaaSHub - Software Alternatives and Reviews. SaaSHub helps you find the best software and product alternatives

NOTE: The number of mentions on this list indicates mentions on common posts plus user suggested alternatives. Hence, a higher number means a better Internet-Places-Database alternative or higher similarity.

Suggest an alternative to Internet-Places-Database

Internet-Places-Database reviews and mentions

Posts with mentions or reviews of Internet-Places-Database. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2024-04-22.

Show HN: OpenOrb, a curated search engine for Atom and RSS feeds
7 projects | news.ycombinator.com | 22 Apr 2024

You can find many RSS feeds, links in my repository
https://github.com/rumca-js/Internet-Places-Database/tree/ma...
It contains also domain lists, that include tag indicating, if it is personal, or not.
We Need to Rewild the Internet
2 projects | news.ycombinator.com | 16 Apr 2024

I am running my personal web crawler since September of 2022. I gather internet domains and assign them meta information. There are various sources of my data. I assign "personal" tag to any personal website. I assign "self-host" tag to any self-host program I find.
I have less than 30k of personal websites.
Data are in the repository.
https://github.com/rumca-js/Internet-Places-Database
I still rely on google for many things, or kagi. It is interesting to me, what my crawler finds next. It is always a surprise to see new blog, or forgotten forum of sorts.
This is how I discover real new content on the Internet. Certainly not by google which can find only BBC, or techcrunch.
The internet is slipping out of our reach
1 project | news.ycombinator.com | 12 Mar 2024

Google will not be interested in fixing search. It also may not be possibile because of ai spam. They would like to invest in deep mind/bard/gemini than to fix technology that will be obsolete in a few years.
I have started scanning domains to see how many different places there are in the internet. Spoiler: Not many.
We could try to create curated open databases for links, forums, places, and links, but in ai era it will always be a niche.
Having said that I think that it is a good thing. If it is a niche it will not be spoiled by normal users expecting simple behavior, or corporations trying to control the output.
Start your blog
Start your curated lists of links.
Control your data. Share your data.
Link https://github.com/rumca-js/Internet-Places-Database
YaCy, a distributed Web Search Engine, based on a peer-to-peer network
9 projects | news.ycombinator.com | 5 Mar 2024

There are already many project about search:
- https://www.marginalia.nu/
- https://searchmysite.net/
- https://lucene.apache.org/
- elastic search
- https://presearch.com/
- https://stract.com/
- https://wiby.me/
I think that all project are fun. I would like to see one succeeding at reaching mainstream level of attention.
I have also been gathering links meta data for some time. Maybe I will use them to feed any eventual self hosted search engine, or language model, if I decide to experiment with that.
- domains for seed https://github.com/rumca-js/Internet-Places-Database
- bookmarks seed https://github.com/rumca-js/RSS-Link-Database
- links for year https://github.com/rumca-js/RSS-Link-Database-2024
A search engine in 80 lines of Python
6 projects | news.ycombinator.com | 7 Feb 2024

I have myself dabbled a little bit in that subject. Some of my notes:
- some RSS feeds are protected by cloudflare. It is true however that it is not necessary for majority of blogs. If you would like to do more then selenium would be a way to solve "cloudflare" protected links
- sometimes even selenium headless is not enough and full blown browser in selenium is necessary to fool it's protection
- sometimes even that is not enough
- then I started to wonder, why some RSS feeds are so well protected by cloudflare, but who am I to judge?
- sometimes it is beneficial to cover user agent. I feel bad for setting my user agent to chrome, but again, why RSS feeds are so well protected?
- you cannot parse, read entire Internet, therefore you always need to think about compromises. For example I have narrowed area of my searches in one of my projects to domains only. Now I can find most of the common domains, and I sort them by their "importance"
- RSS links do change. There need to be automated means to disable some feeds automatically to prevent checking inactive domains
- I do not see any configurable timeout for reading a page, but I am not familiar with aiohttp. Some pages might waste your time
- I hate that some RSS feeds are not configured properly. Some sites do not provide a valid meta "link" with "application/rss+xml". Some RSS feeds have naive titles like "Home", or no title at all. Such a waste of opportunity
My RSS feed parser, link archiver, web crawler: https://github.com/rumca-js/Django-link-archive. Especially interesting could be file rsshistory/webtools.py. It is not advanced programming craft, but it got the job done.
Additionally, in other project I have collected around 2378 of personal sites. I collect domains in https://github.com/rumca-js/Internet-Places-Database/tree/ma... . These files are JSONs. All personal sites have tag "personal".
Most of the things are collected from:
https://nownownow.com/
https://searchmysite.net/
I wanted also to process domains from https://downloads.marginalia.nu/, but haven't got time to read structure of the files
Is Google Getting Worse? A Longitudinal Investigation of SEO Spam in Search [pdf]
6 projects | news.ycombinator.com | 16 Jan 2024

On the other hand it is not 1995. Time has moved on. I wrote a Simple RSS feed, that also serves as search engine for bookmarks.
I am able to run it in attick on raspberry pi. We do not have to rely so heavily on google.
https://github.com/rumca-js/Django-link-archive
It is true that it does not serve me as google, or kagi replacement. It is a very nice addition though.
With a little bit off determination I do not have to be so dependent on google.
Here is also a dump of known domains. Some are personal.
https://github.com/rumca-js/Internet-Places-Database
...and my bookmarks
https://github.com/rumca-js/RSS-Link-Database
Some more years, and google can go to hell.
Ask HN: What apps have you created for your own use?
212 projects | news.ycombinator.com | 12 Dec 2023

[4] https://github.com/rumca-js/Django-link-archive
These are exported then to github repositories:
[5] https://github.com/rumca-js/RSS-Link-Database - bookmarks
[6] https://github.com/rumca-js/RSS-Link-Database-2023 - 2023 year news headlines
[7] https://github.com/rumca-js/Internet-Places-Database - all known to me domains, and RSS feeds
The Small Website Discoverability Crisis
14 projects | news.ycombinator.com | 15 Nov 2023

My own repositories:
- bookmarked entries https://github.com/rumca-js/RSS-Link-Database
- mostly domains https://github.com/rumca-js/Internet-Places-Database
- all 'news' from 2023 https://github.com/rumca-js/RSS-Link-Database-2023
I am using my own Django program to capture and manage links https://github.com/rumca-js/Django-link-archive.
Show HN: List of Internet Domains
1 project | news.ycombinator.com | 30 Oct 2023
A note from our sponsor - SaaSHub
www.saashub.com | 27 Apr 2024

SaaSHub helps you find the best software and product alternatives Learn more →

Stats

Basic Internet-Places-Database repo stats

Mentions

Stars

Activity

9.3

Last Commit

5 days ago

rumca-js/Internet-Places-Database is an open source project licensed under GNU General Public License v3.0 only which is an OSI approved license.

Popular Comparisons