🔍

Google Cloud Professional Data Engineer PDE — Question 58

Topic 1 · Question 58 of 341

Topic 1 · Question 58

You architect a system to analyze seismic data. Your extract, transform, and load (ETL) process runs as a series of MapReduce jobs on an Apache Hadoop cluster. The ETL process takes days to process a data set because some steps are computationally expensive. Then you discover that a sensor calibration step has been omitted. How should you change your ETL process to carry out sensor calibration systematically in the future?