Unlock the power of your data in the cloud! Get hands-on with Google Cloud's core data services like BigQuery and Looker to validate your practical skills in data ingestion, analysis, and management, and earn your Associate Data Practitioner certification!
Prepare and test your skills
Prepare and test your skills
Worked example. The correct answer is already marked and every option is explained below, so there is nothing to select here. To answer questions yourself, start the free trial.
A data analytics team is designing a transformation pipeline to process large volumes of unstructured image data. The pipeline requires intensive linear algebra and numeric matrix computations for feature extraction before the processed metrics are loaded into BigQuery. The team needs a processing engine that supports specialized hardware acceleration (GPUs) to maximize execution performance.
Which Google Cloud product should the team choose to implement this transformation pipeline?
Cloud Dataflow is a fully managed, serverless stream and batch data processing service based on the open-source Apache Beam SDK. It enables large-scale parallel processing of structured, semi-structured, and unstructured data while dynamically provisioning worker compute resources.
When transformation complexity involves unstructured files, heavy mathematical operations, and specialized hardware acceleration requirements, Dataflow is Google Cloud's primary managed processing engine capable of running distributed GPU-accelerated pipelines.
Keep the momentum going with these hand-picked practice scenarios
Want more questions like this?
Get a free certification question every week.