Skip to main page content
U.S. flag

An official website of the United States government

Dot gov

The .gov means it’s official.
Federal government websites often end in .gov or .mil. Before sharing sensitive information, make sure you’re on a federal government site.

Https

The site is secure.
The https:// ensures that you are connecting to the official website and that any information you provide is encrypted and transmitted securely.

Access keys NCBI Homepage MyNCBI Homepage Main Content Main Navigation
. 2000 Jun;25(2):239-40.
doi: 10.1038/76126.

Gene index analysis of the human genome estimates approximately 120,000 genes

Affiliations

Gene index analysis of the human genome estimates approximately 120,000 genes

F Liang et al. Nat Genet. 2000 Jun.

Erratum in

  • Nat Genet 2000 Dec;26(4):501

Abstract

Although sequencing of the human genome will soon be completed, gene identification and annotation remains a challenge. Early estimates suggested that there might be 60,000-100,000 (ref. 1) human genes, but recent analyses of the available data from EST sequencing projects have estimated as few as 45,000 (ref. 2) or as many as 140, 000 (ref. 3) distinct genes. The Chromosome 22 Sequencing Consortium estimated a minimum of 45,000 genes based on their annotation of the complete chromosome, although their data suggests there may be additional genes. The nearly 2,000,000 human ESTs in dbEST provide an important resource for gene identification and genome annotation, but these single-pass sequences must be carefully analysed to remove contaminating sequences, including those from genomic DNA, spurious transcription, and vector and bacterial sequences. We have developed a highly refined and rigorously tested protocol for cleaning, clustering and assembling EST sequences to produce high-fidelity consensus sequences for the represented genes (F.L. et al., manuscript submitted) and used this to create the TIGR Gene Indices-databases of expressed genes for human, mouse, rat and other species (http://www.tigr.org/tdb/tgi.html). Using highly refined and tested algorithms for EST analysis, we have arrived at two independent estimates indicating the human genome contains approximately 120,000 genes.

PubMed Disclaimer

Comment in

  • The nature of the number.
    [No authors listed] [No authors listed] Nat Genet. 2000 Jun;25(2):127-8. doi: 10.1038/75946. Nat Genet. 2000. PMID: 10835616 No abstract available.
  • How to count ... human genes.
    Aparicio SA. Aparicio SA. Nat Genet. 2000 Jun;25(2):129-30. doi: 10.1038/75949. Nat Genet. 2000. PMID: 10835617 No abstract available.

Publication types

LinkOut - more resources