QuestionQ43

Maintaining and automating data workloads

You are operating a Dataflow pipeline that receives messages from a Pub/Sub topic and writes results to a BigQuery dataset in the EU. The pipeline currently runs in europe-west4 with a maximum of 3 n1-standard-1 workers.

During peak periods, the pipeline struggles to process records promptly while all 3 workers have maximum CPU utilization. Which two actions can you take to improve pipeline performance?

Choose two
Explanation

Dataflow horizontal autoscaling can add worker instances up to the configured maximum, increasing parallel processing capacity. Configuring a larger Dataflow worker machine type increases the compute resources available to each worker. These directly address CPU saturation in the worker pool.

Learn more

Community Discussion

No comments yet. Be the first to start the discussion!