I am running a pyspark job on dataproc and my total hdfs capacity is not remaining constant.
As you can see in the first chart that the remaining hdfs capacity is falling even though the used hdfs capacity is minimal. Why is remaining + used not constant?
