In the context of text feature engineering, what does TF-IDF penalize compared to a raw term frequency vector?
-
A
Terms that appear in very few documents
-
B
Terms that appear frequently across many documents
-
C
Terms with very short character length
-
D
Terms that appear only once in a document