You trained a model on data stored in a Cloud Storage bucket. The model needs to be retrained frequently in Vertex AI Training using the latest data in the bucket. Data preprocessing is required prior to retraining. You want to build a simple and efficient near-real-time ML pipeline in Vertex AI that will preprocess the data when new data arrives in the bucket. What should you do?
Cloud Run can be triggered on new data arrivals, which makes it ideal for near-real-time processing. The function then initiates the Vertex AI Pipeline for preprocessing and storing features in Vertex AI Feature Store, aligning with the retraining needs. Cloud Scheduler (Option A) is suitable for scheduled jobs, not event-driven triggers. Dataflow (Option C) is better suited for batch processing or ETL rather than ML preprocessing pipelines.
Shizue
9 months agoFelix
9 months agoDetra
10 months agoRaina
10 months agoLettie
10 months agoRebbeca
10 months agoWinifred
11 months agoEladia
11 months agoMatthew
11 months agoVerona
11 months agoHelga
11 months agoAntonette
11 months agoKeva
11 months agoMike
1 year agoDoug
1 year agoErnest
1 year agoCarri
1 year agoWynell
1 year agoDalene
1 year agoAmie
1 year agoJoesph
1 year agoJolanda
1 year agoCharlesetta
1 year agoCassie
1 year agoVirgina
1 year agoCherrie
1 year agoEve
1 year agoHan
1 year agoTammara
1 year agoTula
2 years agoNieves
2 years agoJavier
1 year agoDallas
1 year agoJackie
2 years ago