files/journal/2022-09-02_11-59-20-000000_418.png

Asian Journal of Information Technology

ISSN: Online 1993-5994
ISSN: Print 1682-3915
121
Views
1
Downloads

Segmenting Broadcast News Streams Using Jlexchains

S. Lalitha and V. Shanthi
Page: 1137-1142 | Received 21 Sep 2022, Published online: 21 Sep 2022

Full Text Reference XML File PDF File

Abstract

In this study, we propose a course-grained NLP approach to text segmentation based on the analysis of lexical cohesion within text. Most research in this area has focused on the discovery of textual units that discuss subtopic structure within documents. In contrast our segmentation task requires the discovery of topical units of text i.e., distinct news stories from broadcast news programmes. Our system SeLeCT first builds a set of lexical chains, in order to model the discourse structure of the text. A boundary detector is then used to search for breaking points in this structure indicated by patterns of cohesive strength and weakness within the text. We evaluate this technique on a test set of concatenated CNN news story transcripts and compare it with an established statistical approach to segmentation called TextTiling.


How to cite this article:

S. Lalitha and V. Shanthi . Segmenting Broadcast News Streams Using Jlexchains.
DOI: https://doi.org/10.36478/ajit.2007.1137.1142
URL: https://www.makhillpublications.co/view-article/1682-3915/ajit.2007.1137.1142