Inductive text classification for medical applications
- 1 January 1995
- journal article
- research article
- Published by Taylor & Francis in Journal of Experimental & Theoretical Artificial Intelligence
- Vol. 7 (1) , 49-80
- https://doi.org/10.1080/09528139508953800
Abstract
Text classification poses a significant challenge for knowledge-based technologies because it touches on all the familiar demons of artificial intelligence: the knowledge engineering bottleneck, problems of scale, easy portability across multiple applications, and cost-effective system construction. Information retrieval (IR) technologies traditionally avoid all of these issues by defining a document in terms of a statistical profile of its lexical items. The IR community is willing to exploit a superficial type of knowledge found in dictionaries and thesaurae, but anything that requires customization, application-specific engineering, or any amount of manual tinkering is thought to be incompatible with practical cost-effective system designs. In this paper those assumptions are challenged and it is shown how machine learning techniques can operate as an effective method for automated knowledge acquisition when it is applied to a representative training corpus, and leveraged against a few hours of routine work by a domain expert. A fully implemented text classification system operating on a medical testbed is described and experimental results based on that testbed are reported.Keywords
This publication has 4 references indexed in Scilit:
- Information Technology Applications in Quality Assurance and Quality improvement, Part IThe Joint Commission Journal on Quality Improvement, 1993
- How Accurate are Hospital Discharge Data for Evaluating Effectiveness of Care?Medical Care, 1993
- Automated Ambulatory Medical Records Systems: An Orphan TechnologyInternational Journal of Technology Assessment in Health Care, 1992
- Knowledge-based Natural Language UnderstandingPublished by Elsevier ,1988