ClickBench Alternatives

Similar projects and alternatives to ClickBench

QuestDB

311 13,420 9.7 Java ClickBench VS QuestDB

An open source time-series database for fast ingest and SQL queries
ClickHouse

208 34,054 10.0 C++ ClickBench VS ClickHouse

ClickHouse® is a free analytics DBMS for big data
InfluxDB

www.influxdata.com sponsored

Power Real-Time Data Analytics at Scale. Get real-time insights from all types of time series data with InfluxDB. Ingest, query, and analyze billions of data points in real-time with unbounded cardinality.
prql

106 9,414 9.9 Rust ClickBench VS prql

PRQL is a modern language for transforming data — a simple, powerful, pipelined SQL replacement
TimescaleDB

82 16,445 9.8 C ClickBench VS TimescaleDB

An open-source time-series SQL database optimized for fast ingest and complex queries. Packaged as a PostgreSQL extension.
Apache Arrow

75 13,442 10.0 C++ ClickBench VS Apache Arrow

Apache Arrow is a multi-language toolbox for accelerated data interchange and in-memory processing
citus

61 9,779 9.5 C ClickBench VS citus

Distributed PostgreSQL as an extension
duckdb

52 16,356 10.0 C++ ClickBench VS duckdb

DuckDB is an in-process SQL OLAP Database Management System
WorkOS

workos.com sponsored

The modern identity platform for B2B SaaS. The APIs are flexible and easy-to-use, supporting authentication, user identity, and complex enterprise features like SSO and SCIM provisioning.
arrow-datafusion

55 4,924 9.9 Rust ClickBench VS arrow-datafusion

Apache DataFusion SQL Query Engine
databend

32 7,157 10.0 Rust ClickBench VS databend

𝗗𝗮𝘁𝗮, 𝗔𝗻𝗮𝗹𝘆𝘁𝗶𝗰𝘀 & 𝗔𝗜. Modern alternative to Snowflake. Cost-effective and simple for massive-scale analytics. https://databend.com
fio

30 4,842 9.3 C ClickBench VS fio

Flexible I/O Tester
uptrace

29 2,874 9.3 Go ClickBench VS uptrace

Open source APM: OpenTelemetry traces, metrics, and logs
hydra

26 2,607 8.7 C ClickBench VS hydra

Hydra: Column-oriented Postgres. Add scalable analytics to your project in minutes. (by hydradatabase)
metaflow

24 7,559 9.2 Python ClickBench VS metaflow

:rocket: Build and manage real-life ML, AI, and data science projects with ease!
chdb

18 1,682 9.6 C++ ClickBench VS chdb

chDB is an embedded OLAP SQL Engine 🚀 powered by ClickHouse
incubation-engineering

18 - - ClickBench VS incubation-engineering
Preql

16 594 0.0 Python ClickBench VS Preql

An interpreted relational query language that compiles to SQL.
pinot

15 5,114 9.9 Java ClickBench VS pinot

Apache Pinot - A realtime distributed OLAP datastore
starrocks

12 7,726 10.0 Java ClickBench VS starrocks

StarRocks, a Linux Foundation project, is a next-generation sub-second MPP OLAP database for full analytics scenarios, including multi-dimensional analytics, real-time analytics, and ad-hoc queries. InfoWorld’s 2023 BOSSIE Award for best open source software.
pg_partman

7 1,863 7.3 PLpgSQL ClickBench VS pg_partman

Partition management extension for PostgreSQL
hosts

306 25,413 9.5 Python ClickBench VS hosts

🔒 Consolidating and extending hosts files from several well-curated sources. Optionally pick extensions for porn, social media, and other categories.
SaaSHub

www.saashub.com sponsored

SaaSHub - Software Alternatives and Reviews. SaaSHub helps you find the best software and product alternatives

NOTE: The number of mentions on this list indicates mentions on common posts plus user suggested alternatives. Hence, a higher number means a better ClickBench alternative or higher similarity.

Suggest an alternative to ClickBench

ClickBench reviews and mentions

Posts with mentions or reviews of ClickBench. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2024-04-16.

Loading a trillion rows of weather data into TimescaleDB
8 projects | news.ycombinator.com | 16 Apr 2024

TimescaleDB primarily serves operational use cases: Developers building products on top of live data, where you are regularly streaming in fresh data, and you often know what many queries look like a priori, because those are powering your live APIs, dashboards, and product experience.
That's different from a data warehouse or many traditional "OLAP" use cases, where you might dump a big dataset statically, and then people will occasionally do ad-hoc queries against it. This is the big weather dataset file sitting on your desktop that you occasionally query while on holidays.
So it's less about "can you store weather data", but what does that use case look like? How are the queries shaped? Are you saving a single dataset for ad-hoc queries across the entire dataset, or continuously streaming in new data, and aging out or de-prioritizing old data?
In most of the products we serve, customers are often interested in recent data in a very granular format ("shallow and wide"), or longer historical queries along a well defined axis ("deep and narrow").
For example, this is where the benefits of TimescaleDB's segmented columnar compression emerges. It optimizes for those queries which are very common in your application, e.g., an IoT application that groups by or selected by deviceID, crypto/fintech analysis based on the ticker symbol, product analytics based on tenantID, etc.
If you look at Clickbench, what most of the queries say are: Scan ALL the data in your database, and GROUP BY one of the 100 columns in the web analytics logs.
- https://github.com/ClickHouse/ClickBench/blob/main/clickhous...
There are almost no time-predicates in the benchmark that Clickhouse created, but perhaps that is not surprising given it was designed for ad-hoc weblog analytics at Yandex.
So yes, Timescale serves many products today that use weather data, but has made different choices than Clickhouse (or things like DuckDB, pg_analytics, etc) to serve those more operational use cases.
Variant in Apache Doris 2.1.0: a new data type 8 times faster than JSON for semi-structured data analysis
2 projects | dev.to | 27 Mar 2024

We tested with 43 Clickbench SQL queries. Queries on the Variant columns are about 10% slower than those on pre-defined static columns, and 8 times faster than those on JSON columns. (For I/O reasons, most cold runs on JSONB data failed with OOM.)
Fair Benchmarking Considered Difficult (2018) [pdf]
2 projects | news.ycombinator.com | 10 Mar 2024

I have a project dedicated to this topic: https://github.com/ClickHouse/ClickBench
It is important to explain the limitations of a benchmark, provide a methodology, and make it reproducible. It also has to be simple enough, otherwise it will not be realistic to include a large number of participants.
I'm also collecting all database benchmarks I could find: https://github.com/ClickHouse/ClickHouse/issues/22398
ClickBench – A Benchmark for Analytical DBMS
1 project | news.ycombinator.com | 8 Feb 2024
FLaNK Stack 05 Feb 2024
49 projects | dev.to | 5 Feb 2024
Why Postgres RDS didn't work for us
4 projects | news.ycombinator.com | 3 Feb 2024

Indeed, ClickHouse results were run on an older instance type of the same family and size (c5.4xlarge for ClickHouse and c6a.4xlarge for Timescale), so if anything ClickHouse results are at a slight disadvantage.
This is an open source benchmark - we'd love contributions from Timescale enthusiasts if we missed something: https://github.com/ClickHouse/ClickBench/
Show HN: Stanchion – Column-oriented tables in SQLite
3 projects | news.ycombinator.com | 31 Jan 2024

Interesting project! Thank you for open sourcing and sharing. Agree that local and embedded analytics are an increasing trend, I see it too.
A couple of questions:
* I’m curious what the difficulties were in the implementation. I suspect it is quite a challenge to implement this support in the current SQLite architecture, and would curious to know which parts were tricky and any design trade-off you were faced with.
* Aside from ease-of-use (install extension, no need for a separate analytical database system), I wonder if there are additional benefits users can anticipate resulting from a single system architecture vs running an embedded OLAP store like DuckDB or clickhouse-local / chdb side-by-side with SQLite? Do you anticipate performance or resource efficiency gains, for instance?
* I am also curious, what the main difficulty with bringing in a separate analytical database is, assuming it natively integrates with SQLite. I may be biased, but I doubt anything can approach the performance of native column-oriented systems, so I'm curious what the tipping point might be for using this extension vs using an embedded OLAP store in practice.
Btw, would love for you or someone in the community to benchmark Stanchion in ClickBench and submit results! (https://github.com/ClickHouse/ClickBench/)
Disclaimer: I work on ClickHouse.
ClickBench: A Benchmark for Analytical Databases
1 project | news.ycombinator.com | 22 Jan 2024
DuckDB performance improvements with the latest release
8 projects | news.ycombinator.com | 6 Nov 2023
DoorDash manages high-availability CockroachDB clusters at scale
1 project | news.ycombinator.com | 2 Nov 2023

interesting. curious if anyone has benchmarked it relative to other dbs. like: https://benchmark.clickhouse.com/
A note from our sponsor - WorkOS
workos.com | 19 Apr 2024

The APIs are flexible and easy-to-use, supporting authentication, user identity, and complex enterprise features like SSO and SCIM provisioning. Learn more →

Stats

Basic ClickBench repo stats

Mentions

Stars

567

Activity

9.1

Last Commit

3 days ago

ClickHouse/ClickBench is an open source project licensed under GNU General Public License v3.0 or later which is an OSI approved license.

The primary programming language of ClickBench is HTML.

ClickBench

ClickBench Alternatives

Similar projects and alternatives to ClickBench

ClickBench reviews and mentions

Stats

Popular Comparisons