Identifying novel transcripts and novel genes in the human genome by using novel SAGE tags
- 4 September 2002
- journal article
- research article
- Published by Proceedings of the National Academy of Sciences in Proceedings of the National Academy of Sciences
- Vol. 99 (19) , 12257-12262
- https://doi.org/10.1073/pnas.192436499
Abstract
The number of genes in the human genome is still a controversial issue. Whereas most of the genes in the human genome are said to have been physically or computationally identified, many short cDNA sequences identified as tags by use of serial analysis of gene expression (SAGE) do not match these genes. By performing experimental verification of more than 1,000 SAGE tags and analyzing 4,285,923 SAGE tags of human origin in the current SAGE database, we examined the nature of the unmatched SAGE tags. Our study shows that most of the unmatched SAGE tags are truly novel SAGE tags that originated from novel transcripts not yet identified in the human genome, including alternatively spliced transcripts from known genes and potential novel genes. Our study indicates that by using novel SAGE tags as probes, we should be able to identify efficiently many novel transcripts/novel genes in the human genome that are difficult to identify by conventional methods.Keywords
This publication has 35 references indexed in Scilit:
- Large-Scale Transcriptional Activity in Chromosomes 21 and 22Science, 2002
- Using the transcriptome to annotate the genomeNature Biotechnology, 2002
- Human Gene Count on the RiseScience, 2002
- A Whale of a Chain ReactionScience, 2002
- The Sequence of the Human GenomeScience, 2001
- Initial sequencing and analysis of the human genomeNature, 2001
- New opportunities for uncovering the molecular basis of cancerNature Genetics, 1997
- Generation and analysis of 280,000 human expressed sequence tags.Genome Research, 1996
- Normalization and subtraction: two approaches to facilitate gene discovery.Genome Research, 1996
- Serial Analysis of Gene ExpressionScience, 1995