Genbank accession
QEI25194.1 [GenBank]
Protein name
tail protein
RBP type
TSP
Evidence RBPdetect
Probability 0,85
TF
Evidence Phold
Probability 1,00
Protein sequence
MKKILDSAREYLRTNNKIKTACLITLELPSSTGSSSAYIYLTDYFRDVIYNGILYTSGKVKSITTHKQNRKLSIGSLSFTITGTAEDEVLKLVQNGVSFLDRSISIHQAVIDEEGNILPVDPDTNGPLLYFRGKITGGGIKDNIGTSGVSTSIITWNCSNQFYDFDRVNGRFTDDADHRGLEIVAGQLVPSNGAKRPEYQEDYGFFHANKSISILAKYQVQEERYKLQSKKKLFGLSRSYSLKKYYETVTKEVDLDFNLAAKFLPVVYGVQKIPGIPVFADTELHNPNIVYVVYAFCEGEIDGFLDFYFGDVPMICIDPNDSKSRTCFGTKKVAGDTMQRIASGQPSSEPSVHGQEYKYNDGNGDIRIWTYHGKADQSAADVLVNIAKEKGFYLQNMNENGPEYWDSRYKLLDTAYAIVRFTINENRTEIPEVSAEVQGKKIKIYHSDGRVTADKTSLNGIWQTLDYLTSDRYGANITIDQFPLQQLIQEAAILDIIDESYQVSWQPYWRYVGWTDAVGENRQIVQMNTILDTSESVFKNVQGLIESYGGAINNLSGQYRITVEKFSNTPLEIDFLDTYGDLELSDTTGRNKFNSVQASILDPALSWKTNSITFFNSIFKEQDKGLDKKLQLSFANITNYYTARSFADRELKKSRYSRTLTFSLPYHFIGIEPNDPIAFTYGRYGWNKKYFLVDEVENSREGKINITLQEYGEDVFINSEQVDNSGNDIPDVSNNVLPPRDFMYTPTPGGQVGSIGKNGELSWLPSLTNNVVYYSIVHSGHADPYIVQQLETNPNLRMIQEIIGEPAGLAVFEIRAVDINGRRSSPVTLSVELNSAKNLSVVSNFRVTNTASGDASEFVGPDVKLAWDRIPEEDIIDGIFYTLEIYDNLDRLLRSVRIEDQYVYDYLLIYNKADYALHNEDALGINRKLRFRIRAEGDNGEQSVDWASI
Physico‐chemical
properties
protein length:949 AA
molecular weight: 107201,68320 Da
isoelectric point:5,05744
aromaticity:0,11275
hydropathy:-0,40263

Domains

Domains [InterPro]
QEI25194.1
1 949
Legend: Pfam SMART CDD TIGRFAM HAMAP SUPFAM PRINTS Gene3D PANTHER Other

Taxonomy

  Name Taxonomy ID Lineage
Phage Salmonella phage SE20
[NCBI]
2592199 Viruses > Duplodnaviria > Heunggongvirae > Uroviricota > Caudoviricetes
Host No host information

Coding sequence (CDS)

Coding sequence (CDS)
Genbank protein accession
QEI25194.1 [NCBI]
Genbank nucleotide accession
MK972709 [NCBI]
CDS location
range 5628 -> 8477
strand -
CDS
ATGAAGAAAATTCTTGACAGTGCAAGAGAGTATTTAAGAACCAATAATAAAATAAAAACTGCATGCCTTATTACTCTTGAATTGCCAAGCTCTACCGGATCAAGTTCTGCGTACATTTATCTTACAGACTATTTTAGGGATGTAATTTATAATGGTATTCTTTACACATCCGGTAAAGTTAAGTCTATAACAACACATAAACAAAATAGAAAACTTTCCATCGGTAGCCTTTCTTTTACTATAACAGGTACTGCTGAGGATGAAGTATTAAAACTAGTCCAAAATGGTGTATCCTTCCTAGATAGATCCATCTCAATCCATCAAGCGGTTATTGATGAAGAAGGTAACATTCTTCCTGTAGATCCAGACACTAATGGTCCTTTACTTTACTTCCGTGGTAAAATCACTGGTGGTGGTATTAAGGACAACATTGGTACTTCTGGAGTTAGTACATCTATTATTACATGGAACTGTTCAAACCAATTCTATGACTTTGATAGAGTAAACGGACGTTTCACAGATGATGCTGACCATAGGGGGCTTGAGATTGTTGCTGGCCAATTAGTTCCCTCTAATGGTGCTAAGCGACCTGAATATCAAGAGGACTACGGATTCTTCCACGCTAACAAGAGTATTTCTATTCTTGCTAAGTATCAGGTTCAGGAAGAACGTTATAAGTTACAATCTAAGAAAAAGTTATTTGGTCTTTCTAGAAGCTATAGCCTAAAGAAATACTATGAGACTGTAACTAAGGAAGTCGATCTTGATTTTAACTTAGCAGCTAAATTCCTTCCAGTAGTATACGGAGTTCAGAAAATTCCTGGAATCCCAGTATTTGCAGATACTGAACTACATAACCCTAATATTGTATATGTAGTATACGCTTTCTGTGAGGGTGAGATTGATGGATTTCTAGATTTCTACTTTGGTGATGTACCAATGATCTGTATTGATCCTAATGATAGTAAGTCTCGTACTTGCTTTGGTACTAAGAAAGTAGCTGGTGATACCATGCAACGTATTGCTTCTGGTCAACCAAGCTCTGAACCTTCTGTTCATGGGCAGGAATATAAGTATAATGATGGTAATGGTGATATCCGTATCTGGACTTATCATGGTAAGGCTGATCAATCTGCTGCTGATGTTCTAGTTAACATTGCAAAAGAGAAAGGTTTCTATTTACAGAACATGAATGAGAATGGACCGGAATACTGGGACTCTCGTTATAAACTGTTAGATACTGCTTATGCTATAGTTCGCTTTACTATTAATGAAAACAGAACAGAAATTCCAGAAGTTAGTGCAGAAGTTCAAGGTAAGAAGATTAAGATCTATCATTCAGATGGTAGAGTAACTGCTGATAAGACTAGCTTAAATGGTATCTGGCAAACTCTGGACTACCTAACCTCTGATCGCTATGGTGCTAATATCACTATCGATCAGTTCCCGCTCCAGCAACTTATTCAAGAAGCAGCTATTTTGGATATCATTGATGAATCCTATCAGGTTTCTTGGCAACCCTATTGGAGATACGTTGGATGGACTGATGCTGTAGGAGAGAATAGACAAATAGTCCAAATGAACACTATTCTGGATACCTCTGAGTCAGTATTTAAAAATGTTCAGGGATTAATAGAGTCTTATGGTGGTGCTATCAACAACTTATCGGGACAGTATAGGATAACTGTTGAGAAGTTCTCAAATACTCCACTAGAGATTGACTTCTTAGATACCTACGGTGATTTGGAGCTATCAGATACCACAGGTAGAAATAAGTTCAACTCAGTACAAGCCTCCATATTAGATCCTGCATTAAGCTGGAAGACTAACTCCATTACATTCTTTAACTCTATATTTAAGGAACAAGATAAAGGATTAGATAAAAAACTTCAGCTTTCTTTTGCAAATATCACTAACTACTATACTGCACGTAGCTTTGCAGATAGAGAACTGAAAAAGTCTCGTTACTCACGAACTCTCACATTCTCATTACCATACCATTTCATAGGCATTGAGCCTAACGATCCGATTGCTTTTACATATGGTCGTTATGGTTGGAATAAGAAATACTTCCTAGTTGATGAGGTTGAAAACTCTAGAGAAGGTAAGATTAATATTACACTGCAAGAGTATGGTGAGGATGTATTCATTAACTCCGAGCAGGTTGATAATAGTGGTAACGATATCCCTGATGTTAGTAATAACGTACTGCCTCCTAGAGACTTTATGTACACACCAACACCGGGTGGACAAGTAGGATCTATTGGTAAGAATGGTGAGCTATCTTGGCTTCCTAGCTTAACTAATAACGTTGTTTATTACTCCATCGTTCATTCTGGACATGCTGATCCTTATATCGTACAGCAACTAGAAACTAATCCAAATCTACGAATGATCCAAGAAATTATTGGAGAACCTGCTGGTTTAGCAGTCTTTGAGATTAGAGCTGTAGATATTAACGGTAGACGAAGTTCTCCTGTAACATTATCGGTAGAACTTAACTCTGCTAAAAACTTAAGCGTCGTTAGTAACTTCAGGGTTACTAATACAGCTTCAGGAGATGCATCGGAGTTTGTAGGACCAGACGTTAAGTTGGCCTGGGACAGAATTCCAGAAGAAGATATCATTGATGGAATCTTCTACACTCTCGAAATCTACGATAACCTAGATCGTTTATTAAGAAGTGTACGAATTGAAGATCAGTATGTCTACGATTATCTACTGATATACAATAAGGCAGACTATGCTCTTCATAACGAGGATGCTCTAGGTATTAATAGGAAGTTACGCTTCCGTATAAGAGCAGAAGGTGATAATGGCGAGCAATCTGTGGATTGGGCATCTATTTAA

Tertiary structure

PDB ID
da598c8f690d1ae19fb757e2b5ab16b2a6774706c9af0c0f9450ed3c11d7c630
ESMFold
Source ESMFold
Method ESMFold
Resolution 0,8212
Oligomeric State monomer
Model Confidence
Very high
pLDDT > 90
High
90 > pLDDT > 70
Low
70 > pLDDT > 50
Very low
pLDDT < 50

Literature

Title Authors Date PMID Source
Salmonella bacteriophage diversity and host specificity revealed by physiological characterization and whole genome sequencing Fong,K., Tremblay,D., Delaquis,P., Moineau,S., Goodridge,L., Levesque,R. and Wang,S. 2019-09-14 GenBank