You've migrated a Hadoop job from an on-premises cluster to Dataproc and Good Storage. Your Spark job is a complex analytical workload fiat consists of many shuffling operations, and initial data are parquet toes (on average 200-400 MB size each) You see some degradation in performance after the migration to Dataproc so you'd like to optimize for it. Your organization is very cost-sensitive so you'd Idee to continue using Dataproc on preemptibles (with 2 non-preemptibles workers only) for this workload. What should you do?
Junita
10 months agoFausto
10 months agoEdgar
10 months agoYuonne
11 months agoMohammad
11 months agoJosephine
11 months agoVincenza
11 months agoKris
11 months agoMarkus
11 months agoLatia
11 months agoEileen
11 months agoLing
11 months agoJanessa
11 months ago