Genbank accession
XVB83608.1 [GenBank]
Protein name
tail fiber protein
RBP type
TSP
Evidence DepoScope
Probability 1,00
TF
Evidence GenBank
Probability 1,00
TF
Evidence Phold
Probability 1,00
TF
Evidence RBPdetect
Probability 0,91
TF
Evidence RBPdetect2
Probability 0,94
TF
Evidence UniProt/TrEMBL
Probability 1,00
Protein sequence
MAVGEIQISALPQAALPIDLSDIFHLKQGIEDKRCTLEQLLAPHSSLRNNPHGVTKTQIGLDNVINALQLVAANNLSDITNVDEARANLQIMSSEEVNNLVQQHINDKSNPHNTTKAQVGLSNVQNWTTSNLYNEDADKYATARAVNNLYKAVQASYPVGTIHLSMNPANPSTYLICGGTWELVSKGRALVGYDSDSRPVGSNFGSSSVSLSSNNLPSHSHSIYLTGGGHTHGAAITIDGFDYGNKSTNSFDYGTKATDTTGAHTHSVSGSTNNTGAHTHTVGGHYGGDSIGGKLRVQVYGTEQVSSVAGDHSHTISGSTNTTGNHQHTVAIGAHTHTVAIGAHTHKGTVTLQSSEHTHSGTTGTTGAGQAFSVEQPSFVVYVWQRTA
Physico‐chemical
properties
protein length:388 AA
molecular weight: 40906,21040 Da
isoelectric point:6,31245
aromaticity:0,05412
hydropathy:-0,40103

Domains

Domains [InterPro]
DC_1039
STR
1–388
IPR051934
Unmapped
97–353
IPR053827
ATT
154–278
XVB83608.1
1 388
Architecture
STR
ATT
STR
STR 1-153 | ATT 154-278 | STR 279-388
Legend: ATT STR RBD CBM LEC ENZ CHP LNK TAS TTP UNK Unmapped

Tail Spike Domain Segmentation

Tail Spike Domain Segmentation

This protein has been segmented into three structural domains: N-terminal, central domain, and C-terminal.

Domain Layout
N-terminal
Central
C-terminal
XVB83608.1
1 388
Domain Start End Length (AA) Confidence
N-terminal 1 239 239 0,5896
Central domain 240 377 139 0,0745
C-terminal 378 388 10 0,9972
Legend: N-terminal Central domain C-terminal
3D Structure with Domain Coloring

The structure is colored according to the domain segmentation: N-terminal (blue), Central (green), C-terminal (pink).

Domain Coloring
N-terminal
1-239
Central
240-377
C-terminal
378-388

Taxonomy

  Name Taxonomy ID Lineage
Phage Salmonella phage P219
[NCBI]
3425684 Viruses > Duplodnaviria > Heunggongvirae > Uroviricota > Caudoviricetes
Host Salmonella sp.
[NCBI]
599 cellular organisms > Bacteria > Pseudomonadati > Pseudomonadota > Gammaproteobacteria > Enterobacterales

Coding sequence (CDS)

Coding sequence (CDS)
Genbank protein accession
XVB83608.1 [NCBI]
Genbank nucleotide accession
PV659053.1 [NCBI]
CDS location
range 19263 -> 20429
strand +
CDS
ATGGCAGTAGGTGAAATTCAAATTAGTGCCTTGCCTCAAGCAGCCTTACCAATTGACCTTAGTGATATCTTCCATCTTAAGCAGGGTATTGAGGATAAGAGATGCACTCTTGAGCAACTACTTGCTCCACACTCAAGCCTAAGAAACAACCCTCATGGTGTTACTAAAACACAGATTGGTTTAGATAATGTTATTAACGCTCTTCAGTTGGTTGCTGCAAATAACTTATCAGACATTACTAATGTTGATGAGGCAAGAGCAAATCTACAGATTATGTCTTCAGAAGAGGTTAATAACCTTGTTCAACAGCATATTAATGATAAGAGTAACCCACATAACACAACTAAGGCACAGGTTGGTTTAAGTAATGTTCAGAACTGGACAACATCTAACCTTTATAATGAAGATGCAGATAAGTACGCTACAGCAAGAGCAGTAAATAACTTGTACAAGGCTGTTCAGGCTTCTTATCCAGTAGGTACTATCCATCTCTCTATGAATCCTGCAAACCCTTCTACATATTTAATTTGTGGGGGTACTTGGGAGTTAGTTTCAAAAGGAAGAGCACTTGTAGGTTATGATAGTGATTCTAGGCCAGTTGGTAGTAACTTTGGCTCAAGTAGTGTTAGCTTATCTAGTAACAACCTACCATCACATAGCCATTCAATCTACCTAACTGGTGGTGGACATACTCATGGTGCTGCTATTACTATCGATGGCTTTGATTACGGCAATAAGAGCACAAACAGTTTCGATTATGGTACTAAAGCCACTGACACTACTGGTGCTCACACCCACTCAGTGAGCGGTTCTACTAACAACACTGGTGCTCATACGCATACTGTTGGTGGTCATTATGGAGGTGACTCTATCGGTGGTAAACTACGTGTTCAGGTATATGGTACAGAACAGGTTTCCAGTGTAGCTGGTGACCACTCACACACTATTAGCGGTTCAACAAATACCACAGGCAATCACCAACACACTGTTGCAATTGGTGCTCACACGCATACTGTTGCAATTGGTGCTCATACCCATAAAGGCACAGTAACTTTGCAGTCATCTGAGCATACTCACTCAGGTACTACAGGTACTACAGGTGCTGGTCAAGCATTCAGCGTTGAACAACCATCCTTTGTGGTTTATGTATGGCAAAGAACCGCTTAA

Genome Context

Genome Context

Tertiary structure

PDB ID
713c0850eb2f7a76394c97cae2f80937348b2a815fcac0dde69be07876f10d77
ESMFold
Source ESMFold
Method ESMFold
Resolution 0,7029
Oligomeric State monomer
Model Confidence
Very high
pLDDT > 90
High
90 > pLDDT > 70
Low
70 > pLDDT > 50
Very low
pLDDT < 50