Genbank accession
AXF53680.1 [GenBank]
Protein name
tail fiber protein and host specificity
RBP type
TF
Evidence Phold
Probability 1,00
TF
Evidence RBPdetect2
Probability 0,86
Protein sequence
MQIWIHDKSMRKVCALNNEIPGMLPYTNSQWHLYLEYSTSTFDFTIPKIVNGKLHDDLKYINDQMYVSFYYDNSYHVFYVSQLVENDFNFQVTCNNTNLELAREVARPLADSGGAKSVEWYLRNLELLGFAGLEIGVNEISDRTRTLTFESQSGTKLEQLHSLMNQFDAEFVFRTDLNRDGTLKKIVIDIYQRPDENHHGIGKVRGDVILYYQSGLKGVQVTSDKTQLFNAGVFTGANGVNLDSVEFEEKNELGQVEFYSRKGTSFVFAPLSRERYPSTMNPDSADNWTRKDFQTEYSDVDSLKAYALRTIKQYAYPLMTYTVSVQSSFIENYKDINLGDTVKIIDNNFRGGLALEARVSEMIISFDNPANNSVVFTNFKKLDNKPSDALQQRIDEIVSKSLPYHVEIRTTNGTVFKNGIGRSTVKPILKQGDKIVDATYRFVIDGTIKYSGMTYDMVASEINQPTTLTISAWVDNKEVASEEVTFVNVSDGKQGPKGDRGNDGKTTYFHTAWSYSADGTDRFSTVYPNLNLLVNSSAKTKDGFFKDFDKVENGYGEVTLTGTNTWVVRNMWDGFSIKPRDYKPGDKYTMSMDVMFTSWNLPAGTHLDEFWIGQRYTHGGGLYSWKRICFIELPKDPSKMLNQWIKITQTSTMPPYEDPAINTEAIFMTKFTGPSEGSFTIRIRNPKQEPGETATPYMPSASEATTADYPSFIGHYTDFTQVDSPNPRDYTWSLIRGNDGKDGANGGENLIVNSAFPEDIDGWGFWDESTPNNNLHIATHGFYYNGTKPLFRLDNNTNGVVPASTKRFPVKRNTDYSLNIQIFATGNLKSVDIYFLGRKANETDKESTKVIHLKTHTGSPSTTQTVKWHLTFNSGDCDEGFIRINNNGTTDGKTSTLFFAELDCYEGTTDRAWQASSKDLEEEIDTKADDVLTQAQLNRLNETNSIIKAELDAKASLDTLNQWVEAYQNFVNANNANRAQAEKDLADASARVTKLENDLNDMSERWNFIDSYMAASNEGLVVGKKDKSSSIMFNPNGRISMFSAGNEVMYISKGVINIENGIFSKTIQIGRYREEQDLLNPDRNVIRYVGGA
Physico‐chemical
properties
protein length:1092 AA
molecular weight: 123332,92340 Da
isoelectric point:5,20437
aromaticity:0,11264
hydropathy:-0,54936

Domains

Domains [InterPro]
IPR010572
ENZ
142–381
AXF53680.1
1 1092
Architecture
STR
STR 1-1092
Legend: ATT STR RBD CBM LEC ENZ CHP LNK TAS TTP UNK Unmapped

Taxonomy

  Name Taxonomy ID Lineage
Phage Streptococcus phage 140
[NCBI]
2268616 No lineage information
Host Streptococcus thermophilus
[NCBI]
1308 cellular organisms > Bacteria > Bacillati > Bacillota > Bacilli > Lactobacillales

Coding sequence (CDS)

Coding sequence (CDS)
Genbank protein accession
AXF53680.1 [NCBI]
Genbank nucleotide accession
MH375598.1 [NCBI]
CDS location
range 16365 -> 19643
strand +
CDS
ATGCAAATTTGGATTCATGATAAAAGCATGCGCAAGGTGTGCGCATTGAATAATGAAATTCCCGGAATGTTGCCATATACGAACAGTCAATGGCATTTATATCTTGAATACTCAACAAGTACGTTTGACTTCACAATTCCTAAAATTGTAAATGGCAAGCTACACGATGATTTAAAATACATCAATGACCAAATGTATGTGTCGTTTTATTATGATAATTCCTACCACGTTTTCTATGTATCGCAACTCGTTGAAAATGATTTTAATTTTCAAGTGACTTGTAATAACACCAACCTTGAATTGGCAAGGGAAGTTGCACGACCACTTGCAGATAGTGGCGGTGCCAAAAGTGTTGAGTGGTATCTTCGGAATCTTGAATTGCTTGGTTTTGCAGGACTTGAAATAGGTGTCAATGAAATTTCTGATAGAACAAGAACGCTTACTTTTGAATCACAAAGTGGCACAAAGCTAGAGCAACTTCATAGCTTGATGAATCAATTTGATGCAGAGTTTGTTTTTCGTACAGATTTAAACCGAGATGGTACTTTAAAAAAAATTGTCATTGACATTTACCAACGACCAGATGAAAACCATCACGGAATTGGAAAGGTTCGAGGAGATGTCATTCTCTACTATCAAAGCGGTCTGAAAGGCGTTCAAGTAACTAGTGATAAGACACAACTGTTTAATGCTGGTGTTTTCACTGGTGCAAACGGGGTTAATCTTGATAGCGTTGAGTTTGAAGAAAAAAATGAGTTAGGACAAGTAGAGTTCTATTCTCGAAAGGGCACTAGCTTCGTTTTCGCCCCACTATCAAGGGAACGCTACCCATCTACCATGAATCCAGACAGCGCTGATAACTGGACACGTAAGGATTTTCAAACAGAATACAGTGATGTTGATTCCCTCAAAGCTTATGCCTTGCGTACTATCAAACAGTATGCTTATCCTCTAATGACATATACTGTCAGCGTTCAATCTAGTTTCATTGAAAACTACAAGGATATTAATCTAGGTGACACTGTTAAAATCATCGATAATAATTTTAGAGGTGGTTTAGCCCTCGAAGCGCGTGTATCTGAAATGATTATCAGCTTTGACAATCCTGCGAATAATTCAGTAGTTTTCACCAACTTTAAAAAGTTGGATAATAAACCATCGGATGCCTTGCAACAACGTATCGATGAGATTGTTTCTAAGTCATTGCCATATCATGTTGAGATAAGGACCACAAACGGTACAGTATTTAAAAACGGTATTGGTCGTTCTACCGTTAAACCAATTTTGAAGCAAGGCGATAAAATTGTTGATGCAACTTATCGATTTGTGATTGACGGTACTATTAAATACTCAGGTATGACCTATGATATGGTAGCATCAGAGATTAACCAACCAACCACGCTTACTATCTCAGCGTGGGTAGATAACAAAGAAGTAGCTTCAGAAGAAGTTACTTTTGTAAATGTATCAGATGGTAAACAAGGACCTAAGGGCGATAGAGGTAATGATGGGAAGACTACATATTTTCACACAGCATGGTCTTACAGCGCAGACGGCACTGATAGGTTTAGTACTGTTTATCCAAATTTGAATCTGTTAGTTAATAGTTCAGCTAAAACCAAAGATGGGTTCTTTAAAGACTTCGACAAAGTAGAAAATGGCTACGGAGAAGTTACATTGACGGGAACTAATACATGGGTTGTTAGAAACATGTGGGATGGTTTCTCTATTAAACCTAGAGATTATAAACCCGGCGATAAGTACACAATGAGTATGGACGTTATGTTCACAAGTTGGAATCTCCCTGCTGGAACACATCTTGACGAGTTTTGGATTGGTCAGCGATACACTCATGGCGGAGGATTATACTCATGGAAGCGTATTTGTTTTATTGAATTACCTAAAGACCCTAGTAAAATGCTGAACCAATGGATAAAAATAACACAAACGTCAACGATGCCTCCGTATGAAGACCCCGCTATCAACACAGAAGCAATCTTTATGACTAAATTTACTGGTCCGAGTGAAGGTAGTTTCACGATAAGGATTAGAAATCCAAAACAAGAACCAGGCGAAACCGCCACTCCATACATGCCATCAGCTAGCGAAGCCACAACTGCTGACTATCCAAGTTTCATCGGACACTACACAGACTTTACACAAGTAGATAGTCCTAATCCTCGAGATTACACTTGGAGTCTGATACGAGGAAACGACGGGAAGGATGGAGCAAATGGTGGAGAGAATCTAATTGTTAATTCAGCATTCCCAGAAGATATTGACGGATGGGGTTTTTGGGACGAAAGTACACCTAATAACAATCTTCATATAGCTACACATGGATTTTACTATAACGGAACAAAACCCCTTTTTAGATTAGACAATAACACCAATGGTGTGGTTCCTGCATCAACAAAACGTTTTCCAGTCAAACGCAACACTGATTATTCTCTAAATATTCAGATATTTGCAACCGGAAACCTCAAGAGCGTTGATATCTATTTTCTTGGCAGGAAGGCGAATGAAACTGACAAGGAATCGACTAAAGTGATCCATTTAAAAACACATACAGGTTCACCATCAACCACACAGACGGTTAAATGGCATCTAACATTTAACTCTGGAGATTGCGACGAAGGGTTCATTCGTATTAATAACAATGGCACTACTGACGGTAAAACTTCTACGCTATTTTTTGCAGAACTAGACTGCTATGAGGGAACCACTGACCGAGCGTGGCAAGCGTCGTCGAAAGATTTGGAAGAGGAAATAGACACCAAAGCCGATGATGTCCTAACACAAGCACAACTCAACAGACTGAACGAAACGAACTCTATTATTAAAGCTGAATTAGACGCTAAAGCGTCGCTTGATACACTCAATCAGTGGGTGGAAGCCTATCAAAATTTTGTTAACGCAAACAATGCCAATCGTGCACAAGCTGAAAAAGATTTAGCTGATGCAAGTGCTCGTGTAACTAAACTAGAAAACGACTTAAATGATATGTCAGAACGTTGGAATTTTATCGATAGCTACATGGCAGCATCAAATGAAGGTCTTGTTGTTGGTAAAAAAGATAAATCAAGCTCTATCATGTTCAATCCAAACGGGCGTATCTCAATGTTCTCAGCTGGGAACGAGGTAATGTACATTTCAAAAGGTGTCATCAATATCGAAAACGGTATTTTCTCTAAAACTATCCAAATCGGACGATATCGAGAGGAACAAGATTTATTGAATCCAGACCGTAATGTCATTAGATACGTAGGAGGTGCATAA

Genome Context

Genome Context

Tertiary structure

PDB ID
814a68037a4411e1e887ff15a4ce85639788959614db548d8713df61c015b3a3
ESMFold
Source ESMFold
Method ESMFold
Resolution 0,7487
Oligomeric State monomer
Model Confidence
Very high
pLDDT > 90
High
90 > pLDDT > 70
Low
70 > pLDDT > 50
Very low
pLDDT < 50

Literature

Title Authors Date PMID Source
Genome analysis of virulent Streptococcus thermophilus bacteriophages isolated in Uruguay Achigar,R., Campot,M.P., Tremblay,D., Labrie,S., Pianzzola,M.J. and Moineau,S. 2017-03-06 GenBank