Google Professional Data Engineer Exam - Topic 4 Question 44 Discussion
You are designing a cloud-native historical data processing system to meet the following conditions:The data being analyzed is in CSV, Avro, and PDF formats and will be accessed by multiple analysis tools including Cloud Dataproc, BigQuery, and Compute Engine.A streaming data pipeline stores new data daily.Peformance is not a factor in the solution.The solution design should maximize availability.How should you design data storage for this solution?
D) Store the data in a multi-regional Cloud Storage bucket. Access the data directly using Cloud Dataproc, BigQuery, and Compute Engine.
A) Create a Cloud Dataproc cluster with high availability. Store the data in HDFS, and peform analysis as needed.
B) Store the data in BigQuery. Access the data using the BigQuery Connector or Cloud Dataproc and Compute Engine.
C) Store the data in a regional Cloud Storage bucket. Aceess the bucket directly using Cloud Dataproc, BigQuery, and Compute Engine.
Johnetta
9 months agoUna
9 months agoDion
9 months agoTashia
9 months agoNguyet
9 months agoDominga
10 months agoKandis
10 months agoCarmelina
10 months agoReuben
10 months agoCarlton
10 months agoJanine
10 months agoEvan
10 months agoRaul
10 months ago