Representing Text Chunks
| dc.creator | Sang, Erik F. Tjong Kim | |
| dc.creator | Veenstra, Jorn | |
| dc.date | 1999-07-06 | |
| dc.date.accessioned | 2026-07-25T16:57:00Z | |
| dc.description | Dividing sentences in chunks of words is a useful preprocessing step for parsing, information extraction and information retrieval. (Ramshaw and Marcus, 1995) have introduced a "convenient" data representation for chunking by converting it to a tagging task. In this paper we will examine seven different data representations for the problem of recognizing noun phrase chunks. We will show that the the data representation choice has a minor influence on chunking performance. However, equipped with the most suitable data representation, our memory-based learning chunker was able to improve the best published chunking results for a standard data set. | |
| dc.description | 7 pages | |
| dc.identifier | https://arxiv.org/abs/cs/9907006 | |
| dc.identifier | http://arxiv.org/abs/cs/9907006 | |
| dc.identifier | EACL'99, Bergen | |
| dc.identifier.uri | https://dspace.dare.co.zw/handle/123456789/44437 | |
| dc.subject | Computation and Language | |
| dc.subject | I.2.7 | |
| dc.title | Representing Text Chunks | |
| dc.type | text |