UniRef
收藏资源简介:
The UniProt Reference Clusters (UniRef) provide clustered sets of sequences from the UniProt Knowledgebase (including isoforms) and selected UniParc records in order to obtain complete coverage of the sequence space at several resolutions while hiding redundant sequences (but not their descriptions) from view. Unlike in UniParc, sequence fragments are merged in UniRef: The UniRef100 database combines identical sequences and sub-fragments with 11 or more residues from any organism into a single UniRef entry, displaying the sequence of a representative protein, the accession numbers of all the merged entries and links to the corresponding UniProtKB and UniParc records.
UniProt参考聚类(UniProt Reference Clusters,简称UniRef)整合了来自UniProt知识库(UniProt Knowledgebase,含蛋白质亚型)与经筛选的UniParc记录的序列簇集,旨在以多种分辨率实现序列空间的完整覆盖,同时隐藏冗余序列(但保留其对应的描述信息)。与UniParc不同,UniRef会对序列片段进行合并:UniRef100数据库将来自任意物种、包含11个及以上残基的相同序列及其亚片段整合为单条UniRef条目,展示代表性蛋白质的序列、所有合并条目的登录号,以及指向对应UniProtKB与UniParc记录的链接。




