seafowl vs litestream

Get real-time insights from all types of time series data with InfluxDB. Ingest, query, and analyze billions of data points in real-time with unbounded cardinality.

www.influxdata.com

featured

SaaSHub - Software Alternatives and Reviews

SaaSHub helps you find the best software and product alternatives

www.saashub.com

featured

seafowl		litestream
	Project
11	Mentions	165
355	Stars	10,026
2.5%	Growth	-
9.3	Activity	7.5
6 days ago	Latest Commit	16 days ago
Rust	Language	Go
Apache License 2.0	License	Apache License 2.0

The number of mentions indicates the total number of mentions that we've tracked plus the number of user suggested alternatives.
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.

seafowl

Posts with mentions or reviews of seafowl. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2023-09-06.

Gcsfuse: A user-space file system for interacting with Google Cloud Storage
15 projects | news.ycombinator.com | 6 Sep 2023

In case you're interested in scale-to-zero database hosting, a few months ago I paired gcsfuse with Seafowl [0][1], an early stage open source database written in Rust. Was a lot of fun balancing tradeoffs that are usually not possible with classical databases e.g. Postgres. Thank you gcsfuse contributors.
[0] https://seafowl.io
DuckDB 0.8.0
5 projects | news.ycombinator.com | 17 May 2023

> why someone would start something in a memory unsafe language these days
You might like what we (Splitgraph) are building with Seafowl [0], a new database which is written in Rust and based on Datafusion and delta-rs [1]. It's optimized for running at the edge and responding to queries via HTTP with cache-friendly semantics.
[0] https://seafowl.io
[1] https://www.splitgraph.com/blog/seafowl-delta-storage-layer
We made a newsfeed for tracking new and deleted datasets across 200+ open data portals (and they're all queryable with SQL)
2 projects | /r/datasets | 13 Apr 2023

For example, here's the IPInfo dataset, and here's a some commodities data from Trase which is proxying to their live Postgres database, and powering their interactive dashboard. Also, here's the repository of Socrata metadata powering the newsfeed - we scrape it nightly and then push it to Seafowl, our new open-source database optimized for running cache-friendly queries "at the edge." The code for Open Data Monitor is on GitHub, if you're curious.
Quicker Serverless Postgres Connections
1 project | news.ycombinator.com | 28 Mar 2023

This is basically how we do authentication in the Splitgraph DDN [0], which is kind of like a multi-tenant serverless Postgres.
We implement the Postgres frontend with a forked version of PgBouncer, and we changed the authentication method such that when the user authenticates, we issue them a JWT which we store as a session variable. That session variable has the same security properties as a cookie in a web browser (the user can change/manipulate it, but if it's signed by us we can trust its claims).
That's the simple explanation that skips over the multi-tenant part. I don't want to derail from the thread - Neon is very cool, and we are actually experimenting with it right now, for storing the Seafowl [1] catalog when deploying to "scale to zero" services like Google Cloud Run or AWS Lambda, which don't have persistent storage.
[0] https://www.splitgraph.com/connect/query
[1] https://seafowl.io
Show HN: Free IP to Country and ASN Downloads from Ipinfo.io
1 project | news.ycombinator.com | 1 Mar 2023

This is really cool! I've always found IP data to be a compelling example of a data product, especially when talking about Splitgraph, a company of which I'm a co-founder (and btw - I also met my co-founder on HN!).
So, I exported the CSV files for country and asn data, and then uploaded them to Splitgraph. You can see some sample queries in the readme of the repository [0]. Since Splitgraph is built on Postgres, it's possible to use all the `inet` and `cidr` tools available from Postgres, so you can make range queries easily. One sample query also demonstrates a join between the two tables, resulting in the equivalent of your combined country_asn.csv.
Another idea: We have a newer project called Seafowl [1], which is an open-source analytical database optimized for running "at the edge," with cache-friendly semantics making it ideal for querying from Web applications. We don't have a self-hosted version of this yet, but perhaps the next thing to try would be loading this data into Seafowl and querying it "at the edge" - I've been thinking about ways that we could package Seafowl along as an OpenResty module, which could allow for true "at the edge" use cases like querying IP data in your reverse proxy. (Although the .mmdb format already solves this particular problem pretty efficiently and interoperably, although I'd be curious to measure the difference).
[0] https://www.splitgraph.com/miles/ipinfo-country-asn
[1] https://seafowl.io/
I Migrated from a Postgres Cluster to Distributed SQLite with LiteFS
4 projects | news.ycombinator.com | 5 Jan 2023

You can indeed run LiteFS by yourself, without Consul, as a sidecar / wrapper around your application. We do it in our project and have a Docker Compose example at [0]. In this case, you specify a specific known leader node. We haven't tried getting it running independently with Consul to do leader election / failover.
[0] https://github.com/splitgraph/seafowl/blob/main/examples/lit...
Ask HN: Serverless SQLite or Closest DX to Cloudflare D1?
2 projects | news.ycombinator.com | 2 Jan 2023

This is the vision of what we're building at Splitgraph. [0] You might be most interested in our recent project Seafowl [1] which is an open-source analytical database optimized for running "at the edge," with cache-friendly semantics making it ideal for querying from Web applications. It's built in Rust using DataFusion and incorporates many of the lessons we've learned building the Data Delivery Network [2] for Splitgraph.
[0] https://www.splitgraph.com
[1] https://seafowl.io
[2] https://www.splitgraph.com/connect
PostgREST – Serve a RESTful API from Any Postgres Database
22 projects | news.ycombinator.com | 29 Dec 2022

> why not just accept SQL and cut out all the unnecessary mapping?
You might be interested in what we're building: Seafowl, a database designed for running analytical SQL queries straight from the user's browser, with HTTP CDN-friendly caching [0]. It's a second iteration of the Splitgraph DDN [1] which we built on top of PostgreSQL (Seafowl is much faster for this use case, since it's based on Apache DataFusion + Parquet).
The tradeoff for allowing the client to run any SQL vs a limited API is that PostgREST-style queries have a fairly predictable and low overhead, but aren't as powerful as fully-fledged SQL with aggregations, joins, window functions and CTEs, which have their uses in interactive dashboards to reduce the amount of data that has to be processed on the client.
There's also ROAPI [2] which is a read-only SQL API that you can deploy in front of a database / other data source (though in case of using databases as a data source, it's only for tables that fit in memory).
[0] https://seafowl.io/
[1] https://www.splitgraph.com/connect
[2] https://github.com/roapi/roapi
Show HN: Socrata Roulette – run random SQL on a random government dataset
1 project | news.ycombinator.com | 9 Dec 2022

It's possible! Currently this is running GROUP BY queries using Socrata's query API on the original government data portal. We're adding the ability to import data from these sources into a columnar format in the future, either into Splitgraph itself or syncing the data out into Seafowl (https://seafowl.io/) which uses Parquet and is much faster.
Technically, the ability is already there (you can add a dataset to Splitgraph and select Socrata as a source if you know the dataset ID), but it's not as turnkey as landing on a dataset page and clicking a button. More to come!
Welcome to InfluxDB IOx: InfluxData’s New Storage Engine
5 projects | news.ycombinator.com | 26 Oct 2022

Just wanted to give a shout out to Apache DataFusion[0] that IOx relies on a lot (and contributes to as well!).
It's a framework for writing query engines in Rust that takes care of a lot of heavy lifting around parsing SQL, type casting, constructing and transforming query plans and optimizing them. It's pluggable, making it easy to write custom data sources, optimizer rules, query nodes etc.
It's has very good single-node performance (there's even a way to compile it with SIMD support) and Ballista [1] extends that to build it into a distributed query engine.
Plenty of other projects use it besides IOx, including VegaFusion, ROAPI, Cube.js's preaggregation store. We're heavily using it to build Seafowl [2], an analytical database that's optimized for running SQL queries directly from the user's browser (caching, CDNs, low latency, some WASM support, all that fun stuff).
[0] https://github.com/apache/arrow-datafusion
[1] https://github.com/apache/arrow-ballista
[2] https://github.com/splitgraph/seafowl

litestream

Posts with mentions or reviews of litestream. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2024-04-07.

Ask HN: SQLite in Production?
3 projects | news.ycombinator.com | 7 Apr 2024

I have not, but I keep meaning to collate everything I've learned into a set of useful defaults just to remind myself what settings I should be enabling and why.
Regarding Litestream, I learned pretty much all I know from their documentation: https://litestream.io/
How (and why) to run SQLite in production
2 projects | news.ycombinator.com | 27 Mar 2024

This presentation is focused on the use-case of vertically scaling a single server and driving everything through that app server, which is running SQLite embedded within your application process.
This is the sweet-spot for SQLite applications, but there have been explorations and advances to running SQLite across a network of app servers. LiteFS (https://fly.io/docs/litefs/), the sibling to Litestream for backups (https://litestream.io), is aimed at precisely this use-case. Similarly, Turso (https://turso.tech) is a new-ish managed database company for running SQLite in a more traditional client-server distribution.
SQLite3 Replication: A Wizard's Guide🧙🏽
2 projects | dev.to | 27 Feb 2024

This post intends to help you setup replication for SQLite using Litestream.
Ask HN: Time travel" into a SQLite database using the WAL files?
1 project | news.ycombinator.com | 2 Feb 2024

I've been messing around with litestream. It is so cool. And, I either found a bug in the -timestamp switch or don't understand it correctly.
What I want to do is time travel into my sqlite database. I'm trying to do some forensics on why my web service returned the wrong data during a production event. Unfortunately, after the event, someone deleted records from the database and I'm unsure what the data looked like and am having trouble recreating the production issue.
Litestream has this great switch: -timestamp. If you use it (AFAICT) you can time travel into your database and go back to the database state at that moment. However, it does not seem to work as I expect it to:
https://github.com/benbjohnson/litestream/issues/564
I have the entirety of the sqlite database from the production event as well. Is there a way I could cycle through the WAL files and restore the database to the point in time before the records I need were deleted?
Will someone take sqlite and compile it into the browser using WASM so I can drag a sqlite database and WAL files into it and then using a timeline slider see all the states of the database over time? :)
Ask HN: Are you using SQLite and Litestream in production?
1 project | news.ycombinator.com | 20 Jan 2024

We're using SQLite in production very heavily with millions of databases and fairly high operations throughput.
But we did run into some scariness around trying to use Litestream that put me off it for the time being. Litestream is really cool but it is also very much a cool hack and the risk of database corruption issues feels very real.
The scariness I ran into was related to this issue https://github.com/benbjohnson/litestream/issues/510
Pocketbase: Open-source back end in 1 file
15 projects | news.ycombinator.com | 6 Jan 2024

Litestream is a library that allows you to easily create backups. You can probably just do analytic queries on the backup data and reduce load on your server.
https://litestream.io/
Litestream – Disaster recovery and continuous replication for SQLite
3 projects | news.ycombinator.com | 1 Jan 2024
Litestream: Replicated SQLite with no main and little cost
1 project | news.ycombinator.com | 6 Nov 2023
Why you should probably be using SQLite
8 projects | news.ycombinator.com | 27 Oct 2023

One possible strategy is to have one directory/file per customer which is one SQLite file. But then as the user logs in, you have to look up first what database they should be connected to.
OR somehow derive it from the user ID/username. Keeping all the customer databases in a single directory/disk and then constantly "lite streaming" to S3.
Because each user is isolated, they'll be writing to their own database. But migrations would be a pain. They will have to be rolled out to each database separately.
One upside is, you can give users the ability to take their data with them, any time. It is just a single file.
[0]. https://litestream.io/
Monitor your Websites and Apps using Uptime Kuma
6 projects | dev.to | 11 Oct 2023

Upstream Kuma uses a local SQLite database to store account data, configuration for services to monitor, notification settings, and more. To make sure that our data is available across redeploys, we will bundle Uptime Kuma with Litestream, a project that implements streaming replication for SQLite databases to a remote object storage provider. Effectively, this allows us to treat the local SQLite database as if it were securely stored in a remote database.

What are some alternatives?

When comparing seafowl and litestream you can also consider the following projects:

marmot - A distributed SQLite replicator built on top of NATS

rqlite - The lightweight, distributed relational database built on SQLite.

datafusion-ballista - Apache Arrow Ballista Distributed Query Engine

pocketbase - Open Source realtime backend in 1 file

azurefs - Mount Microsoft Azure Blob Storage as local filesystem in Linux (inactive)

realtime - Broadcast, Presence, and Postgres Changes via WebSockets

annuaire-entreprises-sirene-api

k8s-mediaserver-operator - Repository for k8s Mediaserver Operator project

mindcastle.io - Massively scalable, cloud-backed distributed block device for Linux and VMs

sqlcipher - SQLCipher is a standalone fork of SQLite that adds 256 bit AES encryption of database files and other security features.

Prisma - Next-generation ORM for Node.js & TypeScript | PostgreSQL, MySQL, MariaDB, SQL Server, SQLite, MongoDB and CockroachDB

litefs - FUSE-based file system for replicating SQLite databases across a cluster of machines

seafowl vs marmot litestream vs rqlite seafowl vs datafusion-ballista litestream vs pocketbase seafowl vs azurefs litestream vs realtime seafowl vs annuaire-entreprises-sirene-api litestream vs k8s-mediaserver-operator seafowl vs mindcastle.io litestream vs sqlcipher seafowl vs Prisma litestream vs litefs

Compare seafowl vs litestream and see what are their differences.

seafowl

litestream

seafowl

litestream

What are some alternatives?