ADAM: another database of abbreviations in MEDLINE
Open Access
- 18 September 2006
- journal article
- research article
- Published by Oxford University Press (OUP) in Bioinformatics
- Vol. 22 (22) , 2813-2818
- https://doi.org/10.1093/bioinformatics/btl480
Abstract
Motivation: Abbreviations are an important type of terminology in the biomedical domain. Although several groups have already created databases of biomedical abbreviations, these are either not public, or are not comprehensive, or focus exclusively on acronym-type abbreviations. We have created another abbreviation database, ADAM, which covers commonly used abbreviations and their definitions (or long-forms) within MEDLINE titles and abstracts, including both acronym and non-acronym abbreviations. Results: A model of recognizing abbreviations and their long-forms from titles and abstracts of MEDLINE (2006 baseline) was employed. After grouping morphological variants, 59 405 abbreviation/long-form pairs were identified. ADAM shows high precision (97.4%) and includes most of the frequently used abbreviations contained in the Unified Medical Language System (UMLS) Lexicon and the Stanford Abbreviation Database. Conversely, one-third of abbreviations in ADAM are novel insofar as they are not included in either database. About 19% of the novel abbreviations are non-acronym-type and these cover at least seven different types of short-form/long-form pairs. Availability: A free, public query interface to ADAM is available at , and the entire database can be downloaded as a text file. Contact: neils@uic.eduKeywords
This publication has 15 references indexed in Scilit:
- Literature mining for the biologist: from information retrieval to biological discoveryNature Reviews Genetics, 2006
- Resolving abbreviations to their senses in MedlineBioinformatics, 2005
- ALICE: An Algorithm to Extract Abbreviations from MEDLINEJournal of the American Medical Informatics Association, 2005
- Biomedical term mapping databasesNucleic Acids Research, 2004
- Achievable Steps Toward Building a National Health Information Infrastructure in the United StatesJournal of the American Medical Informatics Association, 2004
- A Simple and Practical Dictionary-based Approach for Identification of Proteins in Medline AbstractsJournal of the American Medical Informatics Association, 2004
- SaRAD: a Simple and Robust Abbreviation DictionaryBioinformatics, 2004
- Creating an Online Dictionary of Abbreviations from MEDLINEJournal of the American Medical Informatics Association, 2002
- Mapping Abbreviations to Full Forms in Biomedical ArticlesJournal of the American Medical Informatics Association, 2002
- An interactive system for finding complementary literatures: a stimulus to scientific discoveryArtificial Intelligence, 1997