autoloader data processing

Phani1
Databricks MVP

 

Hi Team,

Can you share the best practices for designing the autoloader data processing?

We have data from 30 countries data coming in various files. Currently, we are thinking of using a root folder i.e country, and with subfolders for the individual countries.

In the autoloader script, we plan to set the path to the root folder. Is this a good method? Please advise on the best way to handle thousands of files.

Regards,

Phani