This chapter provides an overview of text clustering, its process, key algorithms, semi-supervised methods, and research from the Beijing Institute of Technology’s NLPIR lab. It covers five major clustering algorithms and their applications, with a focus on text similarity measurement methods like cosine similarity. The chapter also introduces a new method for detecting Top N hot topics using key feature clustering.

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Text Clustering

  • Huaping Zhang,
  • Jianyun Shang

摘要

This chapter provides an overview of text clustering, its process, key algorithms, semi-supervised methods, and research from the Beijing Institute of Technology’s NLPIR lab. It covers five major clustering algorithms and their applications, with a focus on text similarity measurement methods like cosine similarity. The chapter also introduces a new method for detecting Top N hot topics using key feature clustering.