• How many rows does the dataset have?
  • Can you share excel from the display function with the sample?
  • What are the cluster-specific (worker type and runtime version)? Is it standard, high-concurrent, or single-machine?

My blog: https://databrickster.medium.com/