TPEMatcher: A tool for searching in parsed text corpora

Citations

WEB OF SCIENCE

6
Citations

SCOPUS

6

초록

Recently, due to the widespread on-line availability of syntactically annotated text corpora, some automated tools for searching in such text corpora have gained great attention. Generally, those conventional corpus search tools use a decomposition-matching-merging method based on relational predicates for matching a tree pattern query to the desired parts of text corpora. Thus, their query formulation and expressivity are often complicated due to poorly understood query formalisms, and their searching tasks may require a big computational overhead due to a large number of repeated trials of matching tree patterns. To overcome these difficulties, we present TPEMatcher, a tool for searching in parsed text corpora. TPEMatcher provides not only an efficient way of query formulation and searching but also a good query expressivity based on concise syntax and semantics of tree pattern query. We also demonstrate that TPEMatcher can be effectively used for a text mining in practice with its useful interface providing in-depth details of search results.

키워드

Corpus search toolTree pattern queryingTree pattern matchingParsed text corporaText miningANNOTATION
제목
TPEMatcher: A tool for searching in parsed text corpora
저자
Choi, Yong Suk
DOI
10.1016/j.knosys.2011.04.009
발행일
2011-12
유형
Article
저널명
Knowledge-Based Systems
24
8
페이지
1139 ~ 1150