Home LiteratureArticle Details
PMID: 17088284 Published · ppublish English Journal Article Research Support, U.S. Gov't, Non-P.H.S.

The TIGR Plant Transcript Assemblies database.

Nucleic acids research ·Vol. 35 ·No. Database issue ·2007-01-00 ·Pages D846-51

Childs KL, Hamilton JP, Zhu W, Ly E, Cheung F, Wu H, Rabinowicz PD, Town CD, Buell CR, Chan AP

Abstract

The TIGR Plant Transcript Assemblies (TA) database (http://plantta.tigr.org) uses expressed sequences collected from the NCBI GenBank Nucleotide database for the construction of transcript assemblies. The sequences collected include expressed sequence tags (ESTs) and full-length and partial cDNAs, but exclude computationally predicted gene sequences. The TA database includes all plant species for which more than 1000 EST or cDNA sequences are publicly available. The EST and cDNA sequences are first clustered based on an all-versus-all pairwise sequence comparison, followed by the generation of consensus sequences (TAs) from individual clusters. The clustering and assembly procedures use the TGICL tool, Megablast and the CAP3 assembler. The UniProt Reference Clusters (UniRef100) protein database is used as the reference database for the functional annotation of the assemblies. The transcription orientation of each TA is determined based on the orientation of the alignment with the best protein hit. The TA sequences and annotation are available via web interfaces and FTP downloads. Assemblies can be retrieved by a text-based keyword search or a sequence-based BLAST search. The current version of the TA database is Release 2 (July 17, 2006) and includes a total of 215 plant species.

MeSH Terms
DNA, Complementary/chemistry Databases, Nucleic Acid Databases, Protein Expressed Sequence Tags/chemistry Internet Plant Proteins/genetics RNA, Messenger/chemistry RNA, Plant/chemistry User-Computer Interface
Chemicals
DNA, Complementary Plant Proteins RNA, Messenger RNA, Plant
Authors & Affiliations
10 authors, click to expand affiliations / ORCID
Childs Kevin L
The Institute for Genomic Research, 9712 Medical Center Drive, Rockville, MD 20850, USA.
Hamilton John P
Zhu Wei
Ly Eugene
Cheung Foo
Wu Hank
Rabinowicz Pablo D
Town Chris D
Buell C Robin
Chan Agnes P
References (13)
13 references, click to expand
  1. InterPro: an integrated documentation resource for protein families, domains and functional sites.
    Brief Bioinform. 2002 Sep;3(3):225-35 PMID: 12230031
  2. Database resources of the National Center for Biotechnology.
    Nucleic Acids Res. 2003 Jan 1;31(1):28-33 PMID: 12519941
  3. TIGR Gene Indices clustering tools (TGICL): a software system for fast clustering of large EST datasets.
    Bioinformatics. 2003 Mar 22;19(5):651-2 PMID: 12651724
  4. PlantGDB, plant genome database and analysis tools.
    Nucleic Acids Res. 2004 Jan 1;32(Database issue):D354-9 PMID: 14681433
  5. Basic local alignment search tool.
    J Mol Biol. 1990 Oct 5;215(3):403-10 PMID: 2231712
  6. CAP3: A DNA sequence assembly program.
    Genome Res. 1999 Sep;9(9):868-77 PMID: 10508846
  7. The Universal Protein Resource (UniProt).
    Nucleic Acids Res. 2005 Jan 1;33(Database issue):D154-9 PMID: 15608167
  8. The TIGR Gene Indices: clustering and assembling EST and known genes and integration with eukaryotic genomes.
    Nucleic Acids Res. 2005 Jan 1;33(Database issue):D71-4 PMID: 15608288
  9. Comparative plant genomics resources at PlantGDB.
    Plant Physiol. 2005 Oct;139(2):610-8 PMID: 16219921
  10. The Gene Ontology (GO) project in 2006.
    Nucleic Acids Res. 2006 Jan 1;34(Database issue):D322-6 PMID: 16381878
  11. BLAT--the BLAST-like alignment tool.
    Genome Res. 2002 Apr;12(4):656-64 PMID: 11932250
  12. A greedy algorithm for aligning DNA sequences.
    J Comput Biol. 2000 Feb-Apr;7(1-2):203-14 PMID: 10890397
  13. The TIGR Gene Indices: analysis of gene transcript sequences in highly sampled eukaryotic species.
    Nucleic Acids Res. 2001 Jan 1;29(1):159-64 PMID: 11125077
Article Info
Journal
Nucleic acids research
Abbr.
Nucleic Acids Res
ISSN
1362-4962
Published
2007-01-00
Epub
2006-00-06
Pages
D846-51
Language
English
Region
England
NLM ID
0411011
PMCID
PMC1669722
Subset
IM
Analysis Services
Analysis Services

Contact

No. 2 Wenbo Road, Zhangqiu District, Jinan, Shandong

Qilu Normal University · Genelibs Bioinformatics Lab

750 Shunhua Rd, Jinan

2F, Bldg F, University Science Park

Tel: 0531-88819269

WeChat Official Account

Follow our WeChat subscription account for real-time updates and the latest in medical and biological research.


Business Email

E-mail: product@genelibs.com