Academic
Publications
Textractor: A Framework for Extracting Relevant Domain Concepts from Irregular Corporate Textual Datasets

Textractor: A Framework for Extracting Relevant Domain Concepts from Irregular Corporate Textual Datasets,10.1007/978-3-642-12814-1_7,Ashwin Ittoo,Lau

Textractor: A Framework for Extracting Relevant Domain Concepts from Irregular Corporate Textual Datasets   (Citations: 1)
BibTex | RIS | RefWorks Download
Various information extraction (IE) systems for corporate usage exist. However, none of them target the product development and/or customer service domain, despite significant application potentials and benefits. This domain also poses new scientific challenges, such as the lack of external knowledge resources, and irregularities like ungrammatical constructs in textual data, which compromise successful information extraction. To address these issues, we describe the development of Textractor; an application for accurately extracting relevant concepts from irregular textual narratives in datasets of product development and/or customer service organizations. The extracted information can subsequently be fed to a host of business intelligence activities. We present novel algorithms, combining both statistical and linguistic approaches, for the accurate discovery of relevant domain concepts from highly irregular/ungrammatical texts. Evaluations on real-life corporate data revealed that Textractor extracts domain concepts, realized as single or multi-word terms in ungrammatical texts, with high precision.
Conference: Business Information Systems - BIS , pp. 71-82, 2010
Cumulative Annual
View Publication
The following links allow you to view full publications. These links are maintained by other sources not affiliated with Microsoft Academic Search.
Sort by: