Options
- Mark as New
- Bookmark
- Subscribe
- Mute
- Subscribe to RSS Feed
- Permalink
- Report Inappropriate Content
02-08-2022 02:23 AM
Spark will handle the map/reduce for you.
So as long as you use Spark provided functions, be it in scala, python or sql (or even R) you will be using distributed processing.
You just care about what you want as a result.
And afterwards when you are more familiar with Spark you can start tuning (f.e. trying to avoid shuffles, other join types etc)