clusTCR: a Python interface for rapid clustering of large sets of CDR3 sequences
clusTCR: a Python interface for rapid clustering of large sets of CDR3 sequences
复制标题
clusTCR:用于快速聚类大量 CDR3 序列的 Python 接口
DOI:
--
复制
发表时间:
2021
期刊:
影响因子:
--
通讯作者:
P. Meysman
中科院分区:
文献类型:
--
作者:
Sebastiaan Valkiers;Max Van Houcke;K. Laukens;P. Meysman
The T-cell receptor (TCR) determines the specificity of a T-cell towards an epitope. As of yet, the rules for antigen recognition remain largely undetermined. Current methods for grouping TCRs according to their epitope specificity remain limited in performance and scalability. Multiple methodologies have been developed, but all of them fail to efficiently cluster large data sets exceeding 1 million sequences. To account for this limitation, we developed clusTCR, a rapid TCR clustering alternative that efficiently scales up to millions of CDR3 amino acid sequences. Benchmarking comparisons revealed similar accuracy of clusTCR with other TCR clustering methods. clusTCR offers a drastic improvement in clustering speed, which allows clustering of millions of TCR sequences in just a few minutes through efficient similarity searching and sequence hashing. clusTCR was written in Python 3. It is available as an anaconda package (https://anaconda.org/svalkiers/clustcr) and on github (https://github.com/svalkiers/clusTCR).
影响因子:
5.6
作者:
SCHNEIDER, TD;STORMO, GD;EHRENFEUCHT, A
通讯作者:
EHRENFEUCHT, A
影响因子:
5.8
作者:
Sethna, Zachary;Elhanati, Yuval;Mora, Thierry
通讯作者:
Mora, Thierry