sist2
datasette
sist2 | datasette | |
---|---|---|
18 | 187 | |
769 | 8,963 | |
- | - | |
8.5 | 9.3 | |
15 days ago | 4 days ago | |
C | Python | |
GNU General Public License v3.0 only | Apache License 2.0 |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
sist2
-
Better option then filebrowser to share files
Quickly Googling for a docker indexer and search app I turned up Sist2, that on the surface looks like might fit your needs. I don't have an appropriate data store to run it against, so I can't speak to its indexing speed or efficacy. However, the developer does have an accessible demo to try, and the front end at least appears to function well.
-
'google-like' search engine for files on my NAS
I'm also looking for tools like this. You can check out this: https://github.com/simon987/sist2
-
What would you love to see as self hosted service?
Maybe sist2 (https://github.com/simon987/sist2 may fit the bill. It indexes all the metadata and then act as a giant search engine.
- How can I OCR my car manual and make it easy to use in the garage?
- Looking For An App That Will Download Whole Webpages Offline (Specifically Reddit Threads)
-
Seeking a self-hostable search engine for *everything* that I own
I am long user of sist2 from simon987 for full text search of pdf. It indexes everything (file content and metadata) through elasticsearch while providing a nice GUI. https://github.com/simon987/sist2
-
Self hosted web page that indexes all data on a given folder with ability to search? [pi]
I have no experience with this tool, but I recall seeing it in the past. Perhaps it fills your need: https://github.com/simon987/sist2
-
Search engine for local files
sist2 is my primary file indexing / search engine for my SingleFile web archive. Lightweight, blazing fast and tons of customisable options.
-
Docker container with web app for indexing/searching large number of documents
I haven’t tried it in ages but used recoll for local indexing lots of random documents, I found a few repos on GitHub and Docker Hub but nothing super active but may be worth looking at viktor-c/docker-recoll-webui or sist2 is newer and I haven’t used it but may be better maintained at this point
-
Selfhosted File Management Solution? - tags, searching, etc
Having a tool that can scan and index a shared folder would be amazing, and it being accessible from a web browser would also be great, because then I could search from any one of my several devices. The closest thing I have found was sist2. The demo seems to be what I need, but I couldn't seem to get it to run with docker. There's a direct install method, but I haven't tried that yet.
datasette
-
Ask HN: High quality Python scripts or small libraries to learn from
Simon Willison's github would be a great place to get started imo -
https://github.com/simonw/datasette
- Show HN: TextQuery – Query and Visualize Your CSV Data in Minutes
-
Little Data: How do we query personal data? (2013)
I'm a fan on simonw's datasette/dogsheep ecosystem https://datasette.io/
-
LaTeX and Neovim for technical note-taking
I use Anki the exact same way. After a lifetime of learning I have accepted that I will never read over anything I write for myself voluntarily - so my two options are:
1. Write an article so good I can publish it and look it over myself later on. I did this last year with https://andrew-quinn.me/fzf/, for example.
2. Create Anki cards out of the material. Use the builtin Card Browser or even https://datasette.io/ on the underlying SQLite database in a pinch to search for my notes any time I have to.
-
Daily Price Tracking for Trader Joes
Were you aware of, or tempted by https://datasette.io/ for creating your solution?
- SQLite-Web: Web-based SQLite database browser written in Python
-
Ask HN: What two software products should have a kid?
Browsing HN, GitHub and the like we get to see a huge variety of software products and code bases.
I often see products and think - if this product X, got together with Y, it would be pretty cool - kind of like if they had a kid together.
Not too literally, but more on the conceptual level - my level of programming is low.
E.g. Just some....
- pocketable.io & datasette (+with some more charting) [https://pocketbase.io, https://datasette.io]
-
Ask HN: Looking for a project to volunteer on? (February 2024)
You might like the Datasette project: https://datasette.io/
I don't think they are desperate for contributions but it's a welcoming environment and a fun project to hack on. You'll learn a lot just from reading the source and the incredibly informative PRs. The creator is a really talented developer with a great blog which shows up on the HN front page often.
-
Stuff I Learned during Hanukkah of Data 2023
Last year I worked through the challenges using VisiData, Datasette, and Pandas. I walked through my thought process and solutions in a series of posts.
-
What We Watched: A Netflix Engagement Report – About Netflix
> uploads of boring raw excel data and receive a nice UI
https://datasette.io/
What are some alternatives?
Docspell - Assist in organizing your piles of documents, resulting from scanners, e-mails and other sources with miminal effort.
nocodb - 🔥 🔥 🔥 Open Source Airtable Alternative
docker-recoll-webui - Recoll with web frontend and pdf-ocr in a docker container
duckdb - DuckDB is an in-process SQL OLAP Database Management System
Typesense - Open Source alternative to Algolia + Pinecone and an Easier-to-Use alternative to ElasticSearch ⚡ 🔍 ✨ Fast, typo tolerant, in-memory fuzzy Search Engine for building delightful search experiences
sql.js-httpvfs - Hosting read-only SQLite databases on static file hosters like Github Pages
MeiliSearch - A lightning-fast search API that fits effortlessly into your apps, websites, and workflow
litestream - Streaming replication for SQLite.
Ambar - :mag: Ambar: Document Search Engine
Sequel-Ace - MySQL/MariaDB database management for macOS
Gigablast - Nov 20 2017 -- A distributed open source search engine and spider/crawler written in C/C++ for Linux on Intel/AMD. From gigablast dot com, which has binaries for download. See the README.md file at the very bottom of this page for instructions.
beekeeper-studio - Modern and easy to use SQL client for MySQL, Postgres, SQLite, SQL Server, and more. Linux, MacOS, and Windows.