professional-cloud-data-engineer
Prepare and test your skills
Prepare and test your skills
Worked example. The correct answer is already marked and every option is explained below, so there is nothing to select here. To answer questions yourself, start the free trial.
A data engineering team operates a large Apache Spark batch workload on a Google Cloud Dataproc cluster, with input and output datasets stored directly in Cloud Storage buckets. The workload experiences unpredictable processing spikes throughout the day, interspersed with low-utilization periods.
During review of recent cluster operations and billing metrics, the team identifies several performance and cost issues:
Which Dataproc autoscaling policy configuration resolves these issues while minimizing infrastructure costs?
Keep the momentum going with these hand-picked practice scenarios
Want more questions like this?
Get a free certification question every week.