Schema for Highlights - UniProt highlighted "Regions of Interest"
  Database: wuhCor1    Primary Table: unipCov2Interest Data last updated: 2022-10-18
Big Bed File: /gbdb/wuhCor1/uniprot/unipInterestCov2.bb
Item Count: 48
Format description: Browser extensible data (12 fields), eight fields for bigGenePred support, plus extra fields (dbName-pmids, not used by all UniProt subtracks) with UniProt-specific information
fieldexampledescription
chromNC_045512v2Chromosome (or contig, scaffold, etc.)
chromStart21562Start position in chromosome
chromEnd21640End position in chromosome
nameDisorderedName of item
score1000Score from 0-1000
strand++ or -
thickStart21562Start of where display should be thick (start codon)
thickEnd21640End of where display should be thick (stop codon)
reserved12,12,120Used as itemRgb as of 2004-11-22
blockCount1Number of blocks
blockSizes78Comma separated list of block sizes
chromStarts0Start positions relative to chromStart
name2Alternative/human readable name
cdsStartStatcmplStatus of CDS start annotation (none, unknown, incomplete, or complete)
cdsEndStatcmplStatus of CDS end annotation (none, unknown, incomplete, or complete)
exonFrames0Exon frame {0,1,2}, or -1 if no frame for exon
typeswissprotTranscript type
geneNamePrimary identifier for gene
geneName2Alternative/human-readable gene name
geneTypeGene type
statusManually reviewed (Swiss-Prot)Status
annotationTyperegion of interestAnnotation Type
positionamino acids 1-26 on protein P0DTC2Position
longNameLong Name
synsSynonyms
subCellLocSubcell. Location
commentsDisorderedComment
uniProtIdP0DTC2UniProt record
pmids35108439Source articles

Sample Rows
 
chromchromStartchromEndnamescorestrandthickStartthickEndreservedblockCountblockSizeschromStartsname2cdsStartStatcdsEndStatexonFramestypegeneNamegeneName2geneTypestatusannotationTypepositionlongNamesynssubCellLoccommentsuniProtIdpmids
NC_045512v22156221640Disordered1000+215622164012,12,1201780cmplcmpl0swissprotManually reviewed (Swiss-Prot)region of interestamino acids 1-26 on protein P0DTC2DisorderedP0DTC235108439
NC_045512v22176021802Disordered1000+217602180212,12,1201420cmplcmpl0swissprotManually reviewed (Swiss-Prot)region of interestamino acids 67-80 on protein P0DTC2DisorderedP0DTC235108439
NC_045512v22198522054Disordered1000+219852205412,12,1201690cmplcmpl0swissprotManually reviewed (Swiss-Prot)region of interestamino acids 142-164 on protein P0DTC2DisorderedP0DTC235108439
NC_045512v22207822117Disordered1000+220782211712,12,1201390cmplcmpl0swissprotManually reviewed (Swiss-Prot)region of interestamino acids 173-185 on protein P0DTC2DisorderedP0DTC235108439
NC_045512v22229722348Disordered1000+222972234812,12,1201510cmplcmpl0swissprotManually reviewed (Swiss-Prot)region of interestamino acids 246-262 on protein P0DTC2DisorderedP0DTC235108439
NC_045512v22239922465Putative super...1000+223992246512,12,1201660cmplcmpl0swissprotManually reviewed (Swiss-Prot)region of interestamino acids 280-301 on protein P0DTC2Putative superantigen; may bind T-cell receptor alpha/TRACP0DTC232989130
NC_045512v22251623185Receptor-bindi...1000+225162318512,12,12016690cmplcmpl0swissprotManually reviewed (Swiss-Prot)region of interestamino acids 319-541 on protein P0DTC2Receptor-binding domain (RBDP0DTC232132184
NC_045512v22251622567Disordered1000+225162256712,12,1201510cmplcmpl0swissprotManually reviewed (Swiss-Prot)region of interestamino acids 319-335 on protein P0DTC2DisorderedP0DTC235108439
NC_045512v22276822777Integrin-bindi...1000+227682277712,12,120190cmplcmpl0swissprotManually reviewed (Swiss-Prot)region of interestamino acids 403-405 on protein P0DTC2Integrin-binding motif;P0DTC233102950
NC_045512v22287023086binds ACE21000+228702308612,12,12012160cmplcmpl0swissprotManually reviewed (Swiss-Prot)region of interestamino acids 437-508 on protein P0DTC2Receptor-binding motif; binding to human ACE2P0DTC2

Highlights (unipCov2Interest) Track Description
 

Description

This track shows protein sequence annotations defined as "regions of interest" from the UniProt/SwissProt database, mapped to genomic coordinates. The data has been curated from scientific publications by the UniProt/SwissProt staff.

Display Conventions and Configuration

Genomic locations of UniProt/SwissProt annotations are labeled with a short name. A click on the item shows additional annotation deltails.

Mouse-over a feature to see the full UniProt annotation comment.

Methods

UniProt sequences were aligned to UCSC/Gencode transcript sequences first with BLAT, filtered with pslReps (93% query coverage, within top 1% score), lifted to genome positions with pslMap and filtered again. UniProt annotations were obtained from the UniProt XML file. The annotations were then mapped to the genome through the alignment using the pslMap program. This mapping approach draws heavily on the LS-SNP pipeline by Mark Diekhans. Like all Genome Browser source code, the main script used to build this track can be found on GitHub.

Data Access

The raw data can be explored interactively with the Table Browser or the Data Integrator. For automated analysis, the genome annotation is stored in a bigBed file that can be downloaded from the download server. The exact filenames can be found in the track configuration file. Annotations can be converted to ASCII text by our tool bigBedToBed which can be compiled from the source code or downloaded as a precompiled binary for your system. Instructions for downloading source code and binaries can be found here. The tool can also be used to obtain only features within a given range, for example:

bigBedToBed http://hgdownload.soe.ucsc.edu/gbdb/wuhCor1/uniprot/unipInterestCov2.bb -chrom=NC_045512v2 -start=0 -end=29903 stdout

Credits

This track was created by Maximilian Haeussler at UCSC, with help from Chris Lee, Mark Diekhans and Brian Raney, feedback from the UniProt staff and Alejo Mujica, Regeneron Pharmaceuticals. Thanks to UniProt for making all data available for download.

References

UniProt Consortium. Reorganizing the protein space at the Universal Protein Resource (UniProt). Nucleic Acids Res. 2012 Jan;40(Database issue):D71-5. PMID: 22102590; PMC: PMC3245120

Yip YL, Scheib H, Diemand AV, Gattiker A, Famiglietti LM, Gasteiger E, Bairoch A. The Swiss-Prot variant page and the ModSNP database: a resource for sequence and structure information on human protein variants. Hum Mutat. 2004 May;23(5):464-70. PMID: 15108278