News-oriented Automatic Chinese Keyword Indexing

Sujian Li,Houfeng Wang,Shiwen Yu,Chengsheng Xin
DOI: https://doi.org/10.3115/1119250.1119263
2003-01-01
Abstract:In our information era, keywords are very useful to information retrieval, text clustering and so on. News is always a domain attracting a large amount of attention. However, the majority of news articles come without keywords, and indexing them manually costs highly. Aiming at news articles' characteristics and the resources available, this paper introduces a simple procedure to index keywords based on the scoring system. In the process of indexing, we make use of some relatively mature linguistic techniques and tools to filter those meaningless candidate items. Furthermore, according to the hierarchical relations of content words, keywords are not restricted to extracting from text. These methods have improved our system a lot. At last experimental results are given and analyzed, showing that the quality of extracted keywords are satisfying.
What problem does this paper attempt to address?