ML p(r)ior | Multi-Paragraph Segmentation of Expository Text

Multi-Paragraph Segmentation of Expository Text

9406037 | cmp-lg
This paper describes TextTiling, an algorithm for partitioning expository texts into coherent multi-paragraph discourse units which reflect the subtopic structure of the texts. The algorithm uses domain-independent lexical frequency and distribution information to recognize the interactions of multiple simultaneous themes. Two fully-implemented versions of the algorithm are described and shown to produce segmentation that corresponds well to human judgments of the major subtopic boundaries of thirteen lengthy texts.

Highlights - Most important sentences from the article

Login to like/save this paper, take notes and configure your recommendations