About: UCLUST

An Entity of Type: Thing, from Named Graph: http://dbpedia.org, within Data Space: dbpedia.org

UCLUSTis an algorithm designed to cluster nucleotide or amino-acid sequences into clusters based on sequence similarity. The algorithm was published in 2010 and implemented in a program also named UCLUST. The algorithm is described by the author as following two simple clustering criteria, in regard to the requested similarity threshold T. The first criterion states that any given cluster's centroid sequence will have a similarity smaller than T to any other clusters' centroid sequence. The second criterion states that each member sequence in a given cluster will have similarity to the cluster's centroid sequence that is equal or greater than T.

Property	Value
dbo:abstract	UCLUSTis an algorithm designed to cluster nucleotide or amino-acid sequences into clusters based on sequence similarity. The algorithm was published in 2010 and implemented in a program also named UCLUST. The algorithm is described by the author as following two simple clustering criteria, in regard to the requested similarity threshold T. The first criterion states that any given cluster's centroid sequence will have a similarity smaller than T to any other clusters' centroid sequence. The second criterion states that each member sequence in a given cluster will have similarity to the cluster's centroid sequence that is equal or greater than T. UCLUST algorithm is a greedy one. As a result, the order of the sequences in the input file will affect the resulting clusters and their quality. For this reason, it is advised that the sequences will be sorted before entering clustering stage. The program UCLUST is equipped with some options to sort the input sequences prior to clustering them. UCLUST program is widely utilized among the bioinformatic research community, where it used for multiple applications including OTU assignment (e.g. 16s), creating non-redundant gene catalogs, taxonomic assignment and phylogenetic analysis. (en)
dbo:wikiPageExternalLink	http://drive5.com/usearch/manual/uclust_algo.html http://nebc.nerc.ac.uk/bioinformatics/docs/uclust.html
dbo:wikiPageID	45704921 (xsd:integer)
dbo:wikiPageLength	2267 (xsd:nonNegativeInteger)
dbo:wikiPageRevisionID	1079560373 (xsd:integer)
dbo:wikiPageWikiLink	dbr:Nucleotide dbr:Algorithm dbr:Protein_primary_structure dbc:Metagenomics dbr:Gene dbc:2010_software dbr:Taxonomy_(biology) dbr:Bioinformatics dbc:Bioinformatics_algorithms dbr:Sequence_clustering dbr:Phylogenetic_analysis
dbp:wikiPageUsesTemplate	dbt:Cite_web dbt:Reflist dbt:Short_description
dcterms:subject	dbc:Metagenomics dbc:2010_software dbc:Bioinformatics_algorithms
rdfs:comment	UCLUSTis an algorithm designed to cluster nucleotide or amino-acid sequences into clusters based on sequence similarity. The algorithm was published in 2010 and implemented in a program also named UCLUST. The algorithm is described by the author as following two simple clustering criteria, in regard to the requested similarity threshold T. The first criterion states that any given cluster's centroid sequence will have a similarity smaller than T to any other clusters' centroid sequence. The second criterion states that each member sequence in a given cluster will have similarity to the cluster's centroid sequence that is equal or greater than T. (en)
rdfs:label	UCLUST (en)
owl:sameAs	freebase:UCLUST yago-res:UCLUST wikidata:UCLUST https://global.dbpedia.org/id/21gtW
prov:wasDerivedFrom	wikipedia-en:UCLUST?oldid=1079560373&ns=0
foaf:isPrimaryTopicOf	wikipedia-en:UCLUST
is dbo:wikiPageWikiLink of	dbr:QIIME dbr:MG-RAST dbr:Sequence_clustering
is foaf:primaryTopic of	wikipedia-en:UCLUST