Apache Spark with HDFS cluster within Kubernetes
-
Updated
Jul 11, 2023 - Python
Apache Spark with HDFS cluster within Kubernetes
An efficient scheduling system with coflow compression in data-intensive clusters.
Inject errors while running HiBench workloads on Hadoop and collect logs
ansible scripts for setting up multi-cluster hadoop
To associate your repository with the hibench topic, visit your repo's landing page and select "manage topics."