Genbank accession
WQZ52464.1 [GenBank]
Protein name
tail fiber protein
RBP type
TF
Evidence Phold
Probability 1,00
Protein sequence
MKANLKEIENARDKVKSLVGYTYYKRNNLVEDRDTLYSIEAGVIAQDVQTVLPEAVYKIEPQKEDSMLGVSHSGVNALLVNAFNELNEVVEKQQQEIDELKELVKQLLAK
Physico‐chemical
properties
protein length:110 AA
molecular weight: 12493,06720 Da
isoelectric point:4,91404
aromaticity:0,05455
hydropathy:-0,44091

Domains

Domains [InterPro]
IPR030392
CHP
1–97
DC_1202
STR
1–110
IPR030392
CHP
2–57
WQZ52464.1
1 110
Architecture
STR
STR 1-110
Legend: ATT STR RBD CBM LEC ENZ CHP LNK TAS TTP UNK Unmapped

Taxonomy

  Name Taxonomy ID Lineage
Phage Salmonella phage SP33
[NCBI]
3111469 Viruses > Duplodnaviria > Heunggongvirae > Uroviricota > Caudoviricetes
Host Salmonella typhimurium
[NCBI]
90371 cellular organisms > Bacteria > Pseudomonadati > Pseudomonadota > Gammaproteobacteria > Enterobacterales

Coding sequence (CDS)

Coding sequence (CDS)
Genbank protein accession
WQZ52464.1 [NCBI]
Genbank nucleotide accession
OR862218.1 [NCBI]
CDS location
range 75 -> 407
strand +
CDS
ATGAAAGCCAATCTTAAGGAAATTGAGAATGCTCGCGATAAGGTTAAGTCTCTAGTAGGCTATACGTACTATAAGAGAAATAATCTAGTAGAAGACAGAGATACCCTGTACAGTATTGAGGCTGGAGTTATAGCACAAGATGTGCAGACTGTTCTTCCAGAAGCGGTATACAAGATTGAACCACAGAAAGAAGATAGTATGCTTGGTGTGTCTCACTCTGGTGTAAATGCTTTGTTAGTAAATGCTTTTAATGAGCTTAATGAAGTGGTTGAGAAGCAGCAGCAAGAAATTGACGAACTCAAGGAATTGGTAAAACAACTTCTTGCTAAATAA

Genome Context

Genome Context

Tertiary structure

PDB ID
60f8174d2d06de14a83cd392f5b7735551d4114f004f7f37609ad044833080c6
ESMFold
Source ESMFold
Method ESMFold
Resolution 0,8705
Oligomeric State monomer
Model Confidence
Very high
pLDDT > 90
High
90 > pLDDT > 70
Low
70 > pLDDT > 50
Very low
pLDDT < 50