Options
- Mark as New
- Bookmark
- Subscribe
- Mute
- Subscribe to RSS Feed
- Permalink
- Report Inappropriate Content
10-20-2023 11:46 AM
Thanks for your detailed response @Retired_mod!
Regarding the reduce approach, it doesn't seem to work as outlined since reduce is not a member of org.apache.spark.sql.RelationalGroupedDataset
For the second approach, to clarify are you suggesting the following
- Stream silver intermediate table writing with a MERGE INTO statement
- Batch job to read output of (1) and write to gold writing with append
- Stream gold table from (2) to power a different 3rd table