Genbank accession
XJW09878.1 [GenBank]
Protein name
tail fiber protein
RBP type
TF
Evidence GenBank
Probability 1,00
TF
Evidence Phold
Probability 1,00
TSP
Evidence DepoScope
Probability 1,00
TF
Evidence RBPdetect
Probability 0,90
TF
Evidence RBPdetect2
Probability 0,85
Protein sequence
MALKTKIIVQQILNIDDTTTTASKYPKYTVVLGTSISSITASELTGAVEASAASAAAAKDSEIAAKESETNAKDSENLAAIYANSSETSATQSAASATEAERQAGLSKDSADASATSAEESKGFRDSAELAAQNAEQSRLLAEQAKKDAEAAKTAAATSEQNAATSATESTNQAIAAAGSATEAGEYATTAKDSEIAAKTSELNAKNSENESAISAEASEASASQSAISASQSAASATKAAESSAAAKISETTAIESSAAAKTSEINAKTSETNAKTSETNAAAYAAAAKTSETNAADSAASASDSKGFRDEAEAFAAQASTSALAAKNSETNTKTSEINSKASEDAAKLAQQSASGSANTATQAMTTTKGYRDEAEVFKNTATTAATTATDKALEAAGSATIAGEKATNATSAADRAETAAASAEQVMQASLKKDQNLNDLANKDLAREALKVEAVNSVKDQYAGAYNSFRNPAWTYELRIANNGEWRVARNDNNSTSALSIGAGGTGAENVEGAKINFGIDRLKQTETETMMYAPGSNSPYRITIRPDAAWGVWTDETGRWIPLSIDAGGTGSNTEVGARKNLNTPVGGQAIIIPNNSNILGFMSTYAESGYYSSGELVTYQPPEASGWWMYELHVHGKNANGHVEYGNIVATAMNGNKWGIICSAGSWGGWYRIARSDRQLMLLSPDAQSALGDYSIAIGDHDSGLKWDRDGHISAFADSARIFAWTPSGINTYRVISSYVDDNARGMYVNGVRRGDPNALIAGQVEGGSFADWRSRASGLLVEHTGFDSAVAIFKSVYWGKDWIAGMDVVPWTSGGAETHLYVKGAEFIFDSAGNGSASNWVSRSDIRLKAHLKEIETASDKIDYLTGYTYYKRNNLIEDENSVYSIEAGLIAQDVERVLPEAVHSLNNDGQLDPKGEAIKGINYNGVVALLVNAFKEQKAKIDNQQEEINVLRNELYELKNLVKSMLNGNAPTITELP
Physico‐chemical
properties
protein length:983 AA
molecular weight: 103209,52360 Da
isoelectric point:4,82577
aromaticity:0,06409
hydropathy:-0,42960

Domains

Domains [InterPro]
Coil
Unmapped
132–162
IPR030392
CHP
849–909
XJW09878.1
1 983
Architecture
ATT
STR
STR
RBD
ATT 2-193 | STR 252-500 | STR 524-680 | RBD 703-983
Legend: ATT STR RBD CBM LEC ENZ CHP LNK TAS TTP UNK Unmapped

Tail Spike Domain Segmentation

Tail Spike Domain Segmentation

This protein has been segmented into three structural domains: N-terminal, central domain, and C-terminal.

Domain Layout
N-terminal
Central
C-terminal
XJW09878.1
1 983
Domain Start End Length (AA) Confidence
N-terminal 1 366 366 0,7389
Central domain 367 565 200 0,5109
C-terminal 566 983 417 0,7248
Legend: N-terminal Central domain C-terminal
3D Structure with Domain Coloring

The structure is colored according to the domain segmentation: N-terminal (blue), Central (green), C-terminal (pink).

Domain Coloring
N-terminal
1-366
Central
367-565
C-terminal
566-983

Taxonomy

  Name Taxonomy ID Lineage
Phage Salmonella phage SLAM_phiST45
[NCBI]
3373399 Viruses > Duplodnaviria > Heunggongvirae > Uroviricota > Caudoviricetes
Host Salmonella typhimurium
[NCBI]
90371 cellular organisms > Bacteria > Pseudomonadati > Pseudomonadota > Gammaproteobacteria > Enterobacterales

Coding sequence (CDS)

Coding sequence (CDS)
Genbank protein accession
XJW09878.1 [NCBI]
Genbank nucleotide accession
PP948674.1 [NCBI]
CDS location
range 24127 -> 27078
strand +
CDS
ATGGCACTTAAAACTAAAATTATTGTACAGCAGATTCTGAACATAGATGACACTACAACTACTGCTAGTAAATATCCTAAGTATACGGTAGTTTTAGGTACTTCTATTAGTTCTATTACTGCTAGTGAACTAACAGGGGCTGTTGAGGCCTCTGCTGCTTCTGCTGCGGCAGCAAAAGATTCTGAAATTGCAGCAAAAGAATCTGAAACAAATGCTAAGGACTCGGAGAACCTAGCTGCAATTTACGCTAACTCTTCAGAAACTTCTGCAACTCAATCTGCTGCTTCTGCTACTGAAGCGGAGAGACAAGCTGGTTTATCTAAAGATAGTGCTGATGCCTCTGCTACGTCAGCTGAGGAATCCAAAGGATTCCGTGATTCTGCTGAACTTGCTGCACAAAATGCTGAACAGAGTCGTCTATTAGCTGAACAAGCTAAGAAGGACGCTGAGGCTGCTAAGACTGCTGCTGCTACTTCTGAGCAAAATGCTGCTACATCTGCTACTGAGTCTACCAATCAGGCTATTGCTGCTGCTGGTTCTGCTACAGAGGCGGGAGAATACGCTACTACTGCAAAAGACTCCGAGATAGCTGCTAAGACTTCAGAACTTAATGCTAAGAATTCTGAGAATGAATCTGCTATTTCTGCTGAAGCTTCTGAAGCTTCTGCTTCTCAGTCTGCTATTTCTGCTTCTCAATCTGCTGCATCCGCTACTAAAGCTGCAGAATCATCAGCTGCAGCAAAAATTAGTGAAACTACTGCTATAGAATCATCAGCTGCAGCAAAAACTAGTGAGATTAATGCAAAAACTAGTGAGACTAATGCAAAAACTAGTGAGACTAATGCAGCAGCATATGCAGCAGCAGCAAAAACTAGTGAGACTAATGCTGCTGATTCCGCTGCCTCTGCTTCTGACTCCAAAGGATTCAGGGATGAAGCAGAAGCATTCGCTGCACAAGCCTCCACATCAGCATTAGCAGCAAAAAACTCAGAAACTAATACAAAGACTAGTGAAATTAACTCAAAAGCTAGTGAAGACGCTGCTAAGCTAGCTCAGCAAAGTGCATCAGGTAGTGCGAATACAGCTACGCAAGCGATGACCACAACCAAAGGCTACAGAGACGAAGCAGAGGTATTTAAAAATACCGCCACTACTGCTGCAACGACAGCAACAGACAAGGCCCTGGAGGCCGCTGGTAGCGCTACAATAGCAGGAGAGAAAGCTACTAATGCCACCAGTGCAGCAGATAGAGCGGAGACAGCGGCTGCATCAGCAGAACAGGTTATGCAAGCGTCGTTAAAGAAGGATCAGAACCTTAATGATCTTGCAAACAAAGACCTTGCAAGGGAGGCGCTAAAAGTTGAGGCTGTTAATTCTGTAAAAGATCAATATGCTGGCGCTTATAATTCTTTTCGTAATCCTGCATGGACTTATGAATTACGGATCGCTAACAATGGGGAATGGCGCGTTGCGCGTAATGACAATAACAGCACATCTGCGCTTTCTATTGGTGCTGGCGGTACTGGTGCTGAAAACGTTGAGGGCGCAAAAATAAACTTTGGCATTGATCGTTTAAAGCAAACAGAAACAGAAACTATGATGTATGCACCAGGGAGTAACTCACCTTATCGAATCACAATTAGACCTGATGCGGCGTGGGGTGTTTGGACTGATGAAACTGGAAGATGGATTCCTCTTTCAATTGATGCTGGCGGCACTGGATCTAATACTGAAGTTGGAGCAAGAAAGAACTTAAATACTCCTGTTGGTGGTCAAGCAATTATTATTCCTAATAATTCAAACATTCTTGGGTTTATGTCAACATACGCAGAAAGCGGTTATTACTCAAGCGGAGAACTTGTTACATATCAACCACCTGAAGCCTCTGGTTGGTGGATGTATGAATTGCATGTTCACGGTAAAAATGCAAACGGTCATGTTGAGTATGGAAATATTGTTGCAACCGCAATGAACGGTAATAAATGGGGTATCATTTGCAGTGCTGGCTCTTGGGGTGGATGGTACAGGATAGCAAGATCTGATAGGCAATTGATGTTATTAAGTCCGGATGCTCAATCGGCTCTTGGTGATTACTCTATAGCAATTGGGGATCATGATTCAGGTCTTAAATGGGATCGTGATGGTCACATAAGCGCGTTTGCTGATTCAGCGAGAATATTTGCATGGACTCCATCAGGAATAAACACTTACAGGGTCATATCATCTTATGTTGATGATAACGCAAGGGGTATGTATGTAAATGGAGTTAGGCGCGGTGATCCAAACGCTCTTATTGCTGGTCAAGTTGAGGGTGGTTCATTTGCCGACTGGCGTAGTCGTGCGTCTGGATTGCTTGTTGAGCACACTGGATTTGATTCTGCGGTTGCCATATTTAAATCTGTTTATTGGGGGAAAGATTGGATAGCTGGCATGGATGTTGTTCCGTGGACTTCTGGCGGCGCTGAAACACATCTTTATGTTAAGGGCGCTGAATTTATTTTTGATAGTGCTGGAAATGGTAGCGCATCAAATTGGGTTAGCAGGTCTGACATTAGATTAAAAGCACATCTAAAAGAGATCGAAACGGCATCCGACAAAATTGATTATCTAACTGGTTATACTTACTACAAGCGCAACAATCTAATTGAAGATGAAAACAGCGTTTATAGTATTGAGGCTGGATTGATCGCACAAGATGTTGAAAGGGTTTTGCCGGAAGCGGTTCATTCTTTGAATAACGATGGTCAGCTAGACCCAAAAGGCGAGGCAATCAAAGGCATTAACTATAATGGTGTTGTCGCACTTCTTGTTAACGCATTCAAAGAGCAAAAAGCAAAGATTGATAATCAACAGGAGGAAATTAACGTATTACGCAATGAGTTATATGAACTGAAAAATCTTGTAAAATCAATGCTTAACGGAAATGCTCCAACAATTACAGAACTACCGTAA

Genome Context

Genome Context

Tertiary structure

PDB ID
3bdf650bf5f1e7f44536fd66e65883112cdf4afb319e69163f5e3d5477602c10
ESMFold
Source ESMFold
Method ESMFold
Resolution 0,5901
Oligomeric State monomer
Model Confidence
Very high
pLDDT > 90
High
90 > pLDDT > 70
Low
70 > pLDDT > 50
Very low
pLDDT < 50