Avinash_Narala
Databricks Partner

you can make use of databricks native feature "Liquid Clustering", cluster by the columns which you want to use in grouping statements, it will handle the performance issue due to data skewness .

For more information, please do visit :

https://docs.databricks.com/en/delta/clustering.html