You've migrated a Hadoop job from an on-premises cluster to Dataproc and Good Storage. Your Spark job is a complex analytical workload fiat consists of many shuffling operations, and initial data are parquet toes (on average 200-400 MB size each) You see some degradation in performance after the migration to Dataproc so you'd like to optimize for it. Your organization is very cost-sensitive so you'd Idee to continue using Dataproc on preemptibles (with 2 non-preemptibles workers only) for this workload. What should you do?
Junita
9 months agoFausto
9 months agoEdgar
9 months agoYuonne
9 months agoMohammad
9 months agoJosephine
9 months agoVincenza
9 months agoKris
9 months agoMarkus
10 months agoLatia
10 months agoEileen
10 months agoLing
10 months agoJanessa
10 months ago