我正在为一些英语文本生成一些统计数据,我想跳过一些不感兴趣的词,比如"a"和"the".
更新:这些显然被称为"停止词"而不是"跳过词".
language-agnostic indexing nlp filtering stop-words
filtering ×1
indexing ×1
language-agnostic ×1
nlp ×1
stop-words ×1