MCDTB: A Macro-level Chinese Discourse TreeBank

Feng Jiang,Sheng Xu,Xiaomin Chu,Peifeng Li,Qiaoming Zhu,Guodong Zhou
2018-01-01
Abstract:In view of the differences between the annotations of micro and macro discourse rela-tionships, this paper describes the relevant experiments on the construction of the Macro Chinese Discourse Treebank (MCDTB), a higher-level Chinese discourse corpus. Fol-lowing RST (Rhetorical Structure Theory), we annotate the macro discourse information, including discourse structure, nuclearity and relationship, and the additional discourse information, including topic sentences, lead and abstract, to make the macro discourse annotation more objective and accurate. Finally, we annotated 720 articles with a Kappa value greater than 0.6. Preliminary experiments on this corpus verify the computability of MCDTB.
What problem does this paper attempt to address?