Spark Tools
Apache Kafka
Spark Tools | Apache Kafka | |
---|---|---|
- | 26 | |
11 | 27,394 | |
- | 0.8% | |
0.0 | 9.9 | |
10 months ago | about 4 hours ago | |
Scala | Java | |
MIT License | Apache License 2.0 |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
Spark Tools
We haven't tracked posts mentioning Spark Tools yet.
Tracking mentions began in Dec 2020.
Apache Kafka
-
On Implementation of Distributed Protocols
Apache Kafka — a distributed event streaming platform implementing a variant of the Raft consensus protocol (written in Java, integrated with Scala);
- Implementing tagged fields for Kafka Protocol
-
Help me identify this design pattern
Spring does this during autoconfiguration. For example this and this. When the user adds a configuration then it gets to overwrite the default from the template. I am looking for something similar, perhaps simpler approach.
- Kafka Broker Config properties
- Scala DevInTraining looking to contribute to projects
- *bip*
-
What is Kafka ?
Source and documentation on GitHub
-
A simple file source/sink connector?
Code is still in trunk though. https://github.com/apache/kafka/tree/trunk/connect/file/src/main/java/org/apache/kafka/connect/file
-
Can someone please eli5 how the hierarchical timing wheel algorithm works?
I briefly described the algorithm in this article and there is a wonderful article from Kafka that goes into more depth in their general purpose implementation. My implementation is specialized and over optimized in comparison, e.g. by using bit manipulation to avoid more expensive division/modulus instructions. Tokio rewrote their timerwheel after I showed them mine, borrowing some ideas but also staying more general. Hope that helps!
-
How-to-Guide: Contributing to Open Source
Apache Kafka
What are some alternatives?
Sparkta - Real Time Analytics and Data Pipelines based on Spark Streaming
celery - Distributed Task Queue (development branch)
Scoozie - Scala DSL on top of Oozie XML
Apache ActiveMQ Artemis - Mirror of Apache ActiveMQ Artemis
Apache Flink - Apache Flink
redpanda - Redpanda is a streaming data platform for developers. Kafka API compatible. 10x faster. No ZooKeeper. No JVM!
metorikku - A simplified, lightweight ETL Framework based on Apache Spark
jetstream - JetStream Utilities
Scio - A Scala API for Apache Beam and Google Cloud Dataflow.
Aeron - Efficient reliable UDP unicast, UDP multicast, and IPC message transport
Scalding - A Scala API for Cascading
NATS - High-Performance server for NATS.io, the cloud and edge native messaging system.