EasyCluster

EasyCluster produces gene-oriented clusters of Expressed Sequence Tags (ESTs) and Full-Length cDNAs (FL-cDNAs) using genomic sequences to support gene expression and alternative splicing analyses.


Key Features:

  • High Accuracy: Demonstrates superior clustering performance supported by a manually curated benchmark of human EST clusters.
  • Versatile Datasets: Tested on datasets including Unigene cluster Hs.122986 and ESTs from the human HOXA gene family.
  • Comparison with Other Tools: Outperforms genome-based web services such as ASmodeler and BIPASS in clustering capabilities.

Scientific Applications:

  • Gene-Oriented Clustering: Creation of gene-centered clusters of ESTs and FL-cDNAs for gene expression studies.
  • Alternative Splicing Evaluation: Assessment of alternative splicing events from clustered transcripts.
  • Plant Genomics: Compilation of gene-oriented clusters for Ricinus communis, a species lacking Unigene clusters.

Methodology:

Implemented in Python and clusters ESTs and FL-cDNAs by mapping them to genomic sequences to generate gene-oriented clusters.

Topics

Collections

Details

Maturity:
Mature
Cost:
Free of charge
Tool Type:
command-line tool
Operating Systems:
Linux, Mac
Programming Languages:
Python
Added:
4/4/2016
Last Updated:
12/10/2018

Operations

Data Inputs & Outputs

Publications

Picardi E, Mignone F, Pesole G. EasyCluster: a fast and efficient gene-oriented clustering tool for large-scale transcriptome data. BMC Bioinformatics. 2009;10(S6). doi:10.1186/1471-2105-10-s6-s10. PMID:19534735. PMCID:PMC2697633.

Documentation