Training material describes Resilient Distributed Datasets for a flood-sensor analytics job. What core idea should staff take away?
Select an answer to reveal the explanation.
Short Explanation
An RDD is like a sealed, sliced deck of data cards: you do not scribble on the cards in place, and if a slice vanishes, lineage tells Spark how to rebuild it. Mutable free-for-all rows and printed-only sheets miss that resilience story.
Full Explanation
RDDs are immutable, partitioned collections whose resilience comes from lineage that can recompute lost partitions. They are not mutable shared rows for concurrent in-place updates, not merely printed spreadsheets, and they do not replace CAP theorem or eliminate storage needs.