-
incubator-livy
Apache Livy is an open source REST interface for interacting with Apache Spark from anywhere.
-
InfluxDB
Power Real-Time Data Analytics at Scale. Get real-time insights from all types of time series data with InfluxDB. Ingest, query, and analyze billions of data points in real-time with unbounded cardinality.
Spark 3 has some solid support for Kubernetes. Livy is a good addition to Spark on K8S, although support is not merged upstream yet (there's publicly available images and other in the PR). I'd strongly recommend to avoid running HDFS on top of Kubernetes. Either cloud-native buckets or on-premise bucket storage like Minio would be much better suited.
NOTE:
The number of mentions on this list indicates mentions on common posts plus user suggested alternatives.
Hence, a higher number means a more popular project.