An organization is developing a feature repository and is electing to one-hot encode all categorical feature variables. A data scientist suggests that the categorical feature variables should not be one-hot encoded within the feature repository.
Which of the following explanations justifies this suggestion?
In Spark ML, a transformer is an algorithm that can transform one DataFrame into another DataFrame. It takes a DataFrame as input and produces a new DataFrame as output. This transformation can involve adding new columns, modifying existing ones, or applying feature transformations. Examples of transformers in Spark MLlib include feature transformers like StringIndexer, VectorAssembler, and StandardScaler.
Databricks documentation on transformers: Transformers in Spark ML
Marshall
7 months agoRicarda
8 months agoDorthy
8 months agoAlex
8 months agoAsuncion
8 months agoTamar
9 months agoHerman
9 months agoYuonne
9 months agoLelia
9 months agoJesusita
9 months agoFlo
9 months agoLeeann
9 months agoAlida
9 months agoEmily
1 year agoFelice
1 year agoGalen
1 year agoLayla
1 year agoTony
1 year agoHermila
1 year agoIndia
1 year agoJettie
1 year agoStephanie
1 year agoMaurine
1 year agoGaston
1 year agoGlenna
1 year agoChaya
1 year agoMitsue
1 year agoFernanda
1 year agoJoaquin
1 year agoShawna
1 year agoArlette
1 year ago