I have 3 parquet files in ADLS
2 parquet files have 10 sub-parquet files and when i read it as a dataframe in databricks using pyspark, the number of partitions are equal to 10 which is expected behaviour.
3rd file has 172 snappy.parquet files and when i read it as a dataframe, the number of partitions are equal to 89, what is the reason behind this?
Used this command df.rdd.getNumPartitions() to find the number of partitions of a dataframe.