A machine learning engineering team has written predictions computed in a batch job to a Delta table for querying. However, the team has noticed that the querying is running slowly. The team has already tuned the size of the data files. Upon investigating, the team has concluded that the rows meeting the query condition are sparsely located throughout each of the data files.
Based on the scenario, which of the following optimization techniques could speed up the query by colocating similar records while considering values in multiple columns?
Dong
9 months agoKaycee
9 months agoJesse
9 months agoFrederic
10 months agoCarisa
10 months agoShawna
10 months agoLeonida
10 months agoBecky
10 months agoKristel
11 months agoSanjuana
11 months agoPedro
11 months agoAudrie
11 months agoJovita
11 months agoCorazon
11 months agoEmmanuel
11 months agoMelodie
11 months ago