Apache Calcite vs cockroach

Apache Calcite

Apache Calcite (by apache)

Source Code

calcite.apache.org

Suggest alternative

Edit details

cockroach

CockroachDB - the open source, cloud-native distributed SQL database. (by cockroachdb)

Database Go SQL distributed-database Cockroachdb HacktoberFest

Source Code

cockroachlabs.com

Suggest alternative

Edit details

InfluxDB - Power Real-Time Data Analytics at Scale

Get real-time insights from all types of time series data with InfluxDB. Ingest, query, and analyze billions of data points in real-time with unbounded cardinality.

www.influxdata.com

featured

SaaSHub - Software Alternatives and Reviews

SaaSHub helps you find the best software and product alternatives

www.saashub.com

featured

Apache Calcite		cockroach
	Project
28	Mentions	100
4,376	Stars	29,148
1.3%	Growth	0.9%
9.0	Activity	10.0
1 day ago	Latest Commit	1 day ago
Java	Language	Go
Apache License 2.0	License	GNU General Public License v3.0 or later

The number of mentions indicates the total number of mentions that we've tracked plus the number of user suggested alternatives.
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.

Apache Calcite

Posts with mentions or reviews of Apache Calcite. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2023-07-26.

Data diffs: Algorithms for explaining what changed in a dataset (2022)
8 projects | news.ycombinator.com | 26 Jul 2023

> Make diff work on more than just SQLite.
Another way of doing this that I've been wanting to do for a while is to implement the DIFF operator in Apache Calcite[0]. Using Calcite, DIFF could be implemented as rewrite rules to generate the appropriate SQL to be directly executed against the database or the DIFF operator can be implemented outside of the database (which the original paper shows is more efficient).
[0] https://calcite.apache.org/
Apache Baremaps: online maps toolkit
6 projects | news.ycombinator.com | 28 May 2023

Yes, planetiler rocks and the memory mapped collections enabled us to remove our dependency to rocksdb.
From my perspective, planetiler started as an effort to generate vector tiles from the OpenMapTile schema as fast as possible (pbf -> mvt). By contrast, Baremaps started as an effort to create a new schema and style from the ground up. In this regard, having a database (pbf -> db <- mvt) enables to live reload changes made in the configuration files. The database has a cost, but also comes with additional advantages (updates, dynamic data, generation of tiles at zoom levels 16+, etc.).
That being said, I think the two projects overlap and I hope we will find opportunities to collaborate in the future. For instance, whereas PostgreSQL is still required in Baremaps, I recently ported a lot of the ST_ function of Postgis to Apache Calcite with the intent to execute SQL on fast memory mapped collection.
https://github.com/apache/calcite/blob/main/core/src/main/ja...
A planet wide import in Postgis currently takes about 4 hours with the COPY API (easy to parallelize) followed by about 12 hours of simplification in Postgis (not easy to parallelize). I will try to publish a detailed benchmark in the future.
How to manipulate SQL string programmatically?
2 projects | /r/dataengineering | 28 Apr 2023

Use a SQL Parser like sqlglot or Apache Calcite to compile user's query into an AST.
Can SQL be used without an RDBMS?
7 projects | /r/PHP | 27 Feb 2023
Apache Calcite
1 project | news.ycombinator.com | 13 Feb 2023
Want to contribute more to open source projects.
8 projects | /r/dotnet | 18 Aug 2022
CITIC Industrial Cloud — Apache ShardingSphere Enterprise Applications
1 project | dev.to | 14 Apr 2022

The SQL Federation engine contains processes such as SQL Parser, SQL Binder, SQL Optimizer, Data Fetcher and Operator Calculator, suitable for dealing with co-related queries and subqueries cross multiple database instances. At the underlying layer, it uses Calcite to implement RBO (Rule Based Optimizer) and CBO (Cost Based Optimizer) based on relational algebra, and query the results through the optimal execution plan.
Postgres wire compatible SQLite proxy
14 projects | news.ycombinator.com | 31 Mar 2022

Awesome to see work in the DB wire compatible space. On the MySQL side, there was MySQL Proxy (https://github.com/mysql/mysql-proxy), which was scriptable with Lua, with which you could create your own MySQL wire compatible connections. Unfortunately it appears to have been abandoned by Oracle and IIRC doesn't work with 5.7 and beyond. I used it in the past to hack together a MySQL wire adapter for Interana (https://scuba.io/).
I guess these days the best approach for connecting arbitrary data sources to existing drivers, at least for OLAP, is Apache Calcite (https://calcite.apache.org/). Unfortunately that feels a little more involved.
Launch HN: Hydra (YC W22) – Query Any Database via Postgres
4 projects | news.ycombinator.com | 23 Feb 2022

For anyone interested, Apache Calcite[0] is an open source data management framework which seems to do many of the same things that Hydra claims to do, but taking a different approach. Operating as a Java library, Calcite contains "adapters" to many different data sources from existing JDBC connectors to Elasticsearch to Cassandra. All of these different data sources can be joined together as desired. Calcite also has it's own optimizer which is able to push down relevant parts of the query to the different data sources. However, you get full SQL on data sources which don't support it, with Calcite executing the remaining bits itself.
Unfortunately, I would not be too surprised if Calcite was found to be less performance-optimized than Hydra. That said, there are users of Calcite at Google, Uber, Spotify, and others who have made great use of various parts of the framework.
[0] https://calcite.apache.org/
Anyone know of any software that can help in designing then outputting to various database
1 project | /r/DatabaseHelp | 21 Nov 2021

Abstraction Layer - You can use something like Calcite to abstract out your data storage. https://calcite.apache.org/

cockroach

Posts with mentions or reviews of cockroach. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2024-04-11.

11 Planetscale alternatives with free tiers
8 projects | dev.to | 11 Apr 2024

CockroachDB is an open source distributed SQL database designed for scalability and resilience. While it offers SQL databases, CockroachDB is also compatible with PostgreSQL.
A MySQL compatible database engine written in pure Go
10 projects | news.ycombinator.com | 9 Apr 2024

cockroachdb might be close: https://github.com/cockroachdb/cockroach
No More Free Tier on PlanetScale, Here Are Free Alternatives
3 projects | dev.to | 8 Mar 2024

CockroachDB - SQL
Is it bad to create a publicly accessible RDS database for my serverless web app?
2 projects | /r/aws | 11 Aug 2023

For example, when you create a serverless postgres database with a platform like CockroachDB or Neon, you effectively get a connection string with a strong password. Anyone can connect to your database from anywhere so long as they have the right connection string. There are no security settings in these services to change this behavior.
Linux surpasses the Mac among Steam gamers
2 projects | news.ycombinator.com | 4 Aug 2023

> Yes you can on the android emulator. The biggest issue is compu arch in that case.
I can also download VirtualBox and run all Windows programs, that would mean that all Windows apps are Linux apps?
> Yes you can for the most part
You can't statically link glibc: https://github.com/cockroachdb/cockroach/issues/3392
glibc can break stuff: https://www.gamingonlinux.com/2022/08/valve-dev-understandab...
I had binaries break because the newer version if openssl was put under a slightly different name.
How do small SaaS's handle databases?
2 projects | /r/SaaS | 11 Jul 2023

Also, worth noting, if you're already using PostgreSQL (or plan to) you might want to take a look at https://www.cockroachlabs.com/ they have a free tier too and CockroachDB has a PostgreSQL interface.
Go Dependency management in large company projects - How do you do it?
5 projects | /r/golang | 8 Jul 2023

I know that some projects like cockroach use custom build tools like bazel. But we actually really like to use to be able to build our projects simply with the great go toolchain and don't really aim to dive deep into custom build solutions.
Eli5: Why do companies use the products of Oracle to store information, when they can just use spreadsheets like Excel, or make their own spreadsheet software?
1 project | /r/explainlikeimfive | 28 May 2023

CockroachDB is designed to be globally distributed. It has to handle causality when resolving collisions. It has to account for having a write operation to arrive after another and still have time priority because it was sent out a few milliseconds earlier.
rage - a minimalistic load testing tool
2 projects | /r/golang | 27 May 2023

Cockroachdb created a go runtime patch which measures the Grunning time of a goroutine: https://github.com/cockroachdb/cockroach/pull/82356. It doesn't entirely solve the problem though.
Data Engineering Tools in Go
2 projects | /r/dataengineering | 18 May 2023

Our entire backend is written in Go. We've built a platform that allows other companies to offer automatic data syncing to their customers' data warehouses. Go works great for building distributed systems like this (see K8s). We're not the only ones in the space building data intensive applications with Go. Pachyderm, Pinecone, Cockroach Labs and are all also doing it. We've been quite happy with how Go has worked for us.

What are some alternatives?

When comparing Apache Calcite and cockroach you can also consider the following projects:

Trino - Official repository of Trino, the distributed SQL query engine for big data, formerly known as PrestoSQL (https://trino.io)

neon - Neon: Serverless Postgres. We separated storage and compute to offer autoscaling, branching, and bottomless storage.

ANTLR - ANTLR (ANother Tool for Language Recognition) is a powerful parser generator for reading, processing, executing, or translating structured text or binary files.

vitess - Vitess is a database clustering system for horizontal scaling of MySQL.

Presto - The official home of the Presto distributed SQL query engine for big data

tidb - TiDB is an open-source, cloud-native, distributed, MySQL-Compatible database for elastic scale and real-time analytics. Try AI-powered Chat2Query free at : https://tidbcloud.com/free-trial

JSqlParser - JSqlParser parses an SQL statement and translate it into a hierarchy of Java classes. The generated hierarchy can be navigated using the Visitor Pattern

Apache Spark - Apache Spark - A unified analytics engine for large-scale data processing

yugabyte-db - YugabyteDB - the cloud native distributed SQL database for mission-critical applications.

Apache Drill - Apache Drill is a distributed MPP query layer for self describing data

InfluxDB - Scalable datastore for metrics, events, and real-time analytics

Apache Calcite vs Trino cockroach vs neon Apache Calcite vs ANTLR cockroach vs vitess Apache Calcite vs Presto cockroach vs tidb Apache Calcite vs JSqlParser cockroach vs Trino Apache Calcite vs Apache Spark cockroach vs yugabyte-db Apache Calcite vs Apache Drill cockroach vs InfluxDB

Compare Apache Calcite vs cockroach and see what are their differences.

Apache Calcite

cockroach

Apache Calcite

cockroach

What are some alternatives?