Apache Drill
redpanda
Apache Drill | redpanda | |
---|---|---|
9 | 70 | |
1,897 | 8,868 | |
1.1% | 2.5% | |
8.1 | 10.0 | |
6 days ago | 6 days ago | |
Java | C++ | |
Apache License 2.0 | - |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
Apache Drill
-
Git Query Language (GQL) Aggregation Functions, Groups, Alias
Also are you familiar with apache drill . The idea is to put an SQL interpreter in front of any kind of database just like you are doing for git here.
-
Building a Data Lakehouse for Analyzing Elon Musk Tweets using MinIO, Apache Airflow, Apache Drill and Apache Superset
💡 You ca read more here.
- 【上海报案信息】高频关键词 (OC)
-
DeWitt Clause, or Can You Benchmark %DATABASE% and Get Away With It
Apache Drill, Druid, Flink, Hive, Kafka, Spark
- 说的都是什么傻逼东西,一堆警察等了一小时不敢进去还他妈怪拜登不让学校有持枪警卫?哪个财团大得过控制整个共和党的NRA?
-
Apache Drill: the reports of my death have been greatly exaggerated
>We’ve started talking about speeding up our release cadence to better reflect our recent activity.
There's been only one release per year in the past so you can't fault anyone to think the project is dead.
https://github.com/apache/drill/releases
- Concept: A open source alternative to big query ?
-
Roapi: An API Server for Static Datasets
Looks super interesting and potentially useful. Curious how it compares with Apache Drill (https://drill.apache.org/).
-
Does Java have an open source package that can execute SQL on txt/csv?
Check out Apache Drill: https://drill.apache.org/
redpanda
-
Using Redpanda with OpenTelemetry and Grafana for real-time event monitoring
To learn more about Redpanda and stay up-to-date, see Redpanda's source codes available on GitHub and join the Redpanda Community on Slack with fellow developers and data engineers.
-
Choosing Between a Streaming Database and a Stream Processing Framework in Python
Stream-processing platforms such as Apache Kafka, Apache Pulsar, or Redpanda are specifically engineered to foster event-driven communication in a distributed system and they can be a great choice for developing loosely coupled applications. Stream processing platforms analyze data in motion, offering near-zero latency advantages. For example, consider an alert system for monitoring factory equipment. If a machine's temperature exceeds a certain threshold, a streaming platform can instantly trigger an alert and engineers do timely maintenance.
-
The best WebAssembly runtime may be no runtime at all
Yeah it’s just the stack switching itself that is a handful of cycles, but there is not much more overhead for the full VM switch if you structure your embedding the right way. Code the code is source available if you want to peek at it!
https://github.com/redpanda-data/redpanda/blob/dev/src/v/was...
-
redpanda VS quix-streams - a user suggested alternative
2 projects | 7 Dec 2023
-
Kafka Is Dead, Long Live Kafka
that's a littlebit of a stretch. when you say "no shortage" - outside of redpanda what product exists that actually compete in all deployment modes?
it's a misconception that redpanda is simply a better kafka. the way to think about it is that is a new storage engine, from scratch, that speaks the kafka protocol. similar to all of the pgsql companies in a different space, i.e.: big table pgsql support is not a better postgres, fundamentally different tech. you can read the src and design here: https://github.com/redpanda-data/redpanda. or an electric car is not the same as a combustion engine, but only similar in that they are cars that take you from point a to point b.
-
Real-time Data Processing Pipeline With MongoDB, Kafka, Debezium And RisingWave
Redpanda with the MongoDB Debezium Connector installed. We use Redpanda as a Kafka broker.
- Redpanda
-
Flink CDC / alternatives
And Kafka + Kafka Connect has https://www.confluent.io/ https://aiven.io/ https://upstash.com/ (and not quite Kafka, but protocol-compatible, https://redpanda.com/)
-
The Redpanda Project
There exists a C++ project which was created after Rust the language was available. github.com/redpanda-data/redpanda/
-
SOCKS Proxy Server Architecture for High Concurrency
I suggest you check out io_uring and thread per core architecture. Applications like scylladb and redpanda have thread per core architecture and use io_uring for async io.
What are some alternatives?
Trino - Official repository of Trino, the distributed SQL query engine for big data, formerly known as PrestoSQL (https://trino.io)
Apache Kafka - Mirror of Apache Kafka
Apache Calcite - Apache Calcite
Apache Pulsar - Apache Pulsar - distributed pub-sub messaging system
Presto - The official home of the Presto distributed SQL query engine for big data
NATS - High-Performance server for NATS.io, the cloud and edge native messaging system.
AranoDB - The official ArangoDB Java driver.
jetstream - JetStream Utilities
QueryStream - Build JPA Criteria queries using a Stream-like API
RabbitMQ - Open source RabbitMQ: core server and tier 1 (built-in) plugins
spring-data-jpa-mongodb-expressions - Use the MongoDB query language to query your relational database, typically from frontend.
kafkacat - Generic command line non-JVM Apache Kafka producer and consumer [Moved to: https://github.com/edenhill/kcat]