Torna ai risultati
Scheda bibliografica · Consultazione e accesso
Artículo

Cd-hit: a fast program for clustering and comparing large sets of protein or nucleotide sequences

Weizhong Li; Adam Godzik · Bioinformatics · 2006

Pagina della risorsa
Lettura rapida. Controlla i dati essenziali della risorsa e accedi al contenuto con il pulsante principale. La scheda mostra solo le informazioni necessarie per identificare, citare e aprire l’opera.

Accesso alla risorsa

Apri il contenuto dall’opzione principale o scegli un’altra fonte disponibile.

OpenAlex OpenAlex Works
Entrar por OpenAlex
Accesso principale

Pagina della risorsa

Pagina di riferimento della risorsa. La disponibilità del testo completo non è stata confermata automaticamente.
Apri risorsa

Riepilogo

Descripción general del contenido del recurso.

MOTIVATION: In 2001 and 2002, we published two papers (Bioinformatics, 17, 282-283, Bioinformatics, 18, 77-82) describing an ultrafast protein sequence clustering program called cd-hit. This program can efficiently cluster a huge protein database with millions of sequences. However, the applications of the underlying algorithm are not limited to only protein sequences clustering, here we present several new programs using the same algorithm including cd-hit-2d, cd-hit-est and cd-hit-est-2d. Cd-hit-2d compares two protein datasets and reports similar matches between them; cd-hit-est clusters a DNA/RNA sequence database and cd-hit-est-2d compares two nucleotide datasets. All these programs can handle huge datasets with millions of sequences and can be hundreds of times faster than methods based on the popular sequence comparison and database search tools, such as BLAST.

Come citare

Elegí el formato que necesitás y copiá la referencia al portapapeles.

APA 7

Li, W. & Godzik, A. (2006). Cd-hit: a fast program for clustering and comparing large sets of protein or nucleotide sequences. https://doi.org/10.1093/bioinformatics/btl158

MLA

Li, Weizhong, and Adam Godzik. "Cd-hit: a fast program for clustering and comparing large sets of protein or nucleotide sequences." 2006. https://doi.org/10.1093/bioinformatics/btl158.

Chicago

Li, Weizhong and Adam Godzik. 2006. "Cd-hit: a fast program for clustering and comparing large sets of protein or nucleotide sequences.". https://doi.org/10.1093/bioinformatics/btl158.

Harvard

Li, W. and Godzik, A. 2006, Cd-hit: a fast program for clustering and comparing large sets of protein or nucleotide sequences, Bioinformatics, available at: https://doi.org/10.1093/bioinformatics/btl158 [Accessed 7 Aug. 2026].

Condividi e stampa

Salva la scheda, copia il link permanente o stampala in PDF.

Esporta riferimento

Esporta il record nei formati più comuni per usarlo con un gestore bibliografico.

Dettagli della risorsa

Informazioni bibliografiche utili per verificare che sia il materiale corretto.

Titolo
Cd-hit: a fast program for clustering and comparing large sets of protein or nucleotide sequences
Autore / collaboratori
Weizhong Li; Adam Godzik
Editore
Bioinformatics
Anno di pubblicazione
2006
Lingua
Inglés

Soggetti

Esplora risorse correlate a partire da questi soggetti.

Copiato