Nightly merges of three 200-million-row sources need a repeatable Spark job, job bookmarks, and a catalog update. Which service should run that transform?
Select an answer to reveal the explanation.
Short Explanation
Think of nightly merges of three 200-million-row sources, job bookmarks, and a catalog update. AWS Glue. A laptop notebook and a modest Data Wrangler flow do not swallow that.
Full Explanation
AWS Glue Spark is the scheduled, large-scale lake-transform path: bookmarks and a catalog update included. A laptop notebook and a modest Data Wrangler flow do not merge 200 million rows. Rekognition is not that job.