Best stage order for a Spark ML text classification Pipeline?
Select an answer to reveal the explanation.
Short Explanation and Infographic
Tokenize → vectorize → classify. Estimators come after their feature transformers.
Full explanation below image
Full Explanation
Spark ML Pipelines require feature transformers before classifiers that consume feature vectors.