Protein
View in Explore- Genbank accession
- AZF91682.1 [GenBank]
- Protein name
- tail fiber protein and host specificity
- RBP type
-
TF
- Protein sequence
-
MLLTIHDANLQKIGFIDNEKQETLNFYNDTWTRNLETASSTFEFTVSKKQLLSDTGNKHLYNQLNERSFVSFKYKGKTYLFNIMKTEENERWLRCYCENLNLELINEYTNAYKADRPRSFAEYLDVFEIPQFAMVKIGVNEISDQKRTLEWEGQETKLARLLSLANKFDAEVEFVTQLNDDSSIKQLVLNVYHKADDSHTGVGRIRGDIRLTFEKNIKSMTRKIDKTEVYTLVVPYGKSKETHEGEQEVRVYIDSLPPWEEKNDEGIVIFKQEGVNLYAPHAADLYPSTFGVSTQSNKWIRKDLEVDSDNPSVIRAAGIANLRKHAYPAITYEVDGFVDVEIGDTITIHDKGFTPALDIRSRAVEQKISFSNPTNNKTTFGNFKELENRTSGDLRSVFEQMVENSRPYNILVSTDNGVMFKNNTGRSTISPTLKRGNQTVPATYRFVIDGSIVSSGLTYTVKASDISKPTVITISAWVDNKEVASEEVTFVNVSDGKQGPKGDRGNDGLPGKDGVGLKSTTITYGMSDSDTIMPTSWTANPPTLIKGKYLWTKTQWTYTDSTSETGYQKTYIAKDGNNGNDGLPGKDGVGIRNTTITYAQGTSGTVAPTNGWSTQVPNVPAGQYLWTKTVWDYTDNTSETGYSVAKMGEQGPRGDQGIPGPKGIDGTDAPTIFVKSYTYSAGSKAYIKLTGPNAFEQTLYYSRGHNVWVLDATTHKLKEFVHCDTYITMSFNHNGVNITLADYLNSITDSIVAIAAADADAVDQNFRDVLNKMGGNPELGTWSWRTGHVFIGMSKRSDGTWPLQPRQGYEVAIQEDGSAPEIGCTLSIGGIVANGADGKTQYTHIAYANSADGSKDFSTSDSNRAYIGMYVDFNINDSTNPSDYSWTLVKGADGTQGVPGKPGADGKTPYFHTAWAYSADGTDGFTTVYPNLNLLEGTATFDGMNPNSSDNSVSAITKTKISGIANTVMDVKTSGNAFAVGFYTQKGYNITAGQTITISFIAKASSDTSLFVGFEHFPSGHKMFTISTKWELYTYTFTATTSGTPTFVMYGWDMVAGQGFQLYNPKAELGSVATPWMPSASEVTTADYPSFIGQYTNYTQVDSPNPRDYTWSLIRGNDGKQGPQGPKGDRGIPGIKGADGRTQYTHIAYADTISGSGFSQTDVNKAYIGMYQDFNAEDSKNPQDYRWSKWKGSDGRDGIPGKAGADGRTPYVHFAYADSADGQKGFSLTQTGRKRYLGVLTNFFKEDSTNPSDYTWNDTTGSISVGGRNLLVKTNQGITNWNWQLSDGDKSVEEVKVDGIRAVKLIKGSTAANTGWNFIEYNGLLRELIQPKSKYVLSFDVKPSVDVTFYATLARGDFNEPLTDTVAMPKALANQWNKVSCVLTSKETLPNIAWQVVYLAGMPTTNGNWVIIKNIKLEEGDIPTQWTPAIEDIQDEIDSKADAAMTIEQINALNERAGIIKAEMEAKASAEILNNWIKNYQDFVKANETERAAAEKALVSSSQRVSTIAKELGELSDRWNFIDTYMNSSNDGLVIGKNDGSSSMMFNPNGRISMYSAGEEVMYISQGVIHIENGIFSKTIQVGRYREEQYHLNPDMNVIRYVGGF
- Physico‐chemical
properties -
protein length: 1605 AA molecular weight: 177357,12800 Da isoelectric point: 5,30350 aromaticity: 0,10530 hydropathy: -0,50617
Domains
Domains [InterPro]
Legend:
Pfam
SMART
CDD
TIGRFAM
HAMAP
SUPFAM
PRINTS
Gene3D
PANTHER
Other
Taxonomy
| Name | Taxonomy ID | Lineage | |
|---|---|---|---|
| Phage |
Streptococcus phage CHPC1057 [NCBI] |
2365020 | Viruses > Duplodnaviria > Heunggongvirae > Uroviricota > Caudoviricetes |
| Host | No host information | ||
Coding sequence (CDS)
Coding sequence (CDS)
Genbank protein accession
AZF91682.1
[NCBI]
Genbank nucleotide accession
MH937498.1
[NCBI]
CDS location
range 16249 -> 21066
strand +
strand +
CDS
ATGTTGCTAACAATTCACGACGCTAATTTGCAGAAGATTGGCTTCATCGATAACGAGAAGCAAGAAACGTTAAATTTCTATAATGATACTTGGACCCGAAACCTTGAAACAGCATCAAGCACGTTCGAGTTTACTGTTTCTAAAAAACAGTTACTTAGCGATACAGGAAATAAACACCTTTATAACCAACTAAACGAGCGCTCTTTTGTTTCCTTCAAATATAAGGGCAAGACATATCTTTTTAACATCATGAAGACGGAAGAGAATGAGCGATGGTTACGATGTTATTGCGAAAACTTAAATCTCGAGCTGATAAATGAGTACACGAACGCCTACAAGGCTGACAGGCCTAGGTCTTTCGCGGAATATCTTGATGTATTCGAAATTCCTCAGTTTGCGATGGTTAAAATAGGTGTTAATGAGATCTCTGATCAGAAGAGAACGCTCGAGTGGGAAGGGCAGGAAACAAAACTGGCAAGGCTCTTAAGCTTGGCCAATAAATTTGACGCTGAAGTTGAATTTGTGACTCAACTTAACGATGACAGTTCAATTAAGCAACTCGTTTTGAACGTTTACCATAAAGCGGACGACTCGCACACTGGTGTGGGTCGGATTCGTGGCGATATTCGTCTTACGTTTGAAAAGAATATCAAATCAATGACGAGAAAGATTGATAAGACTGAAGTCTATACGCTGGTAGTTCCTTATGGCAAATCGAAAGAGACCCATGAAGGCGAGCAAGAAGTACGTGTCTACATTGACAGCCTTCCGCCTTGGGAGGAAAAGAACGATGAAGGTATTGTTATCTTCAAACAAGAAGGTGTCAACCTCTATGCACCTCATGCAGCCGACTTATACCCGTCTACTTTTGGCGTATCAACTCAATCTAATAAATGGATTCGGAAAGACCTTGAGGTCGATAGTGACAATCCGAGCGTTATCCGTGCCGCAGGAATTGCGAACTTGCGTAAACACGCCTATCCCGCTATCACTTACGAGGTTGACGGTTTTGTAGATGTAGAAATCGGGGATACCATAACAATCCACGACAAAGGATTCACACCAGCGCTTGACATAAGATCGCGTGCCGTTGAGCAAAAAATCAGCTTCAGCAATCCGACCAACAATAAGACCACTTTCGGAAACTTCAAAGAGCTTGAAAATAGGACGTCTGGAGACCTTAGGAGCGTTTTCGAGCAAATGGTTGAGAACAGTCGACCATACAATATCCTAGTCTCAACTGATAACGGTGTTATGTTTAAGAACAACACAGGGCGGTCAACCATAAGTCCAACGTTGAAACGAGGAAACCAGACCGTTCCCGCAACATATCGCTTTGTAATTGATGGCTCTATTGTTAGCTCTGGTCTGACCTATACCGTCAAAGCAAGCGATATCTCAAAACCAACTGTGATCACGATTTCAGCGTGGGTTGATAACAAAGAAGTAGCTTCAGAAGAAGTTACTTTTGTAAATGTATCAGATGGTAAACAAGGACCTAAGGGCGATAGAGGTAATGACGGCTTACCGGGTAAAGACGGGGTAGGCTTGAAATCTACCACAATCACTTACGGCATGTCTGACAGTGATACTATCATGCCTACGAGTTGGACTGCAAACCCACCAACTTTGATTAAAGGGAAATACCTATGGACTAAAACTCAGTGGACGTATACTGATAGTACAAGTGAGACTGGTTATCAAAAGACTTATATTGCTAAAGATGGTAATAACGGTAATGATGGCCTTCCGGGTAAAGATGGCGTTGGTATACGTAATACCACAATCACTTACGCACAAGGAACATCTGGAACAGTAGCACCAACGAATGGTTGGAGTACTCAAGTGCCGAATGTACCTGCTGGACAATACCTATGGACTAAAACAGTCTGGGATTATACTGACAACACTAGTGAAACTGGATATTCAGTCGCTAAAATGGGCGAGCAAGGTCCAAGAGGTGACCAAGGTATACCTGGTCCTAAAGGTATTGATGGTACTGATGCTCCAACGATTTTCGTTAAGTCCTATACATACTCAGCAGGTTCAAAGGCCTATATTAAACTGACTGGGCCAAATGCTTTTGAGCAAACCTTATATTACAGCCGAGGACACAATGTGTGGGTTCTTGATGCTACAACACATAAACTCAAAGAGTTCGTACATTGTGATACCTATATAACCATGTCATTTAATCATAATGGTGTTAATATAACATTGGCTGACTACCTAAATAGTATTACAGATAGTATTGTCGCAATTGCAGCAGCGGATGCAGACGCAGTTGACCAAAATTTTAGGGATGTGCTTAACAAAATGGGTGGTAATCCAGAACTTGGAACATGGAGTTGGCGAACTGGTCACGTCTTTATAGGCATGTCCAAGCGGTCTGATGGAACCTGGCCACTGCAACCACGACAGGGGTATGAAGTAGCCATACAAGAAGATGGATCAGCACCAGAAATTGGATGCACTCTGTCAATAGGAGGAATAGTTGCTAATGGAGCAGACGGTAAAACACAATATACCCATATTGCATACGCGAATAGCGCAGATGGAAGTAAAGATTTTTCAACTTCTGATTCTAATCGTGCCTATATCGGGATGTACGTTGATTTTAACATCAATGATTCAACCAATCCGAGCGATTACTCATGGACACTTGTTAAAGGTGCTGACGGTACTCAAGGTGTACCAGGAAAACCTGGGGCTGACGGGAAGACTCCTTATTTCCATACGGCGTGGGCTTACAGTGCAGATGGTACCGATGGTTTCACGACTGTTTACCCTAATTTGAATTTGTTGGAAGGGACCGCTACGTTTGACGGAATGAACCCTAACTCTAGTGATAATTCGGTTAGTGCTATTACAAAAACTAAAATATCAGGAATTGCTAATACAGTCATGGACGTAAAAACAAGCGGAAATGCTTTTGCCGTTGGTTTTTATACACAAAAAGGTTATAACATAACCGCTGGGCAGACCATTACTATTTCATTTATAGCAAAAGCATCAAGTGACACAAGTCTTTTTGTTGGATTTGAACATTTTCCAAGTGGACATAAAATGTTCACAATAAGCACAAAGTGGGAACTTTATACTTATACATTCACAGCAACAACGTCAGGAACTCCAACTTTTGTGATGTACGGGTGGGATATGGTAGCAGGGCAAGGATTCCAATTATACAACCCTAAAGCGGAACTAGGTTCAGTTGCTACCCCTTGGATGCCCTCGGCTAGCGAAGTCACAACTGCTGATTATCCAAGTTTCATCGGACAATATACAAACTATACACAAGTAGATAGTCCTAATCCTCGAGATTACACTTGGAGCCTCATTCGAGGTAACGATGGTAAACAAGGACCACAAGGTCCTAAAGGTGACCGAGGGATACCAGGGATAAAAGGTGCTGACGGAAGAACGCAGTATACCCACATAGCTTATGCTGATACAATTTCAGGTAGTGGCTTTAGTCAAACAGATGTCAATAAAGCCTATATTGGTATGTATCAAGACTTCAATGCCGAAGATAGCAAAAATCCACAAGATTATCGTTGGTCTAAGTGGAAAGGTAGTGATGGACGAGATGGTATTCCAGGAAAAGCTGGGGCTGACGGACGTACGCCTTACGTCCATTTTGCTTATGCCGATAGTGCCGATGGTCAAAAAGGTTTCAGTTTGACACAAACTGGACGCAAGCGCTATTTAGGTGTGCTTACCAACTTCTTCAAGGAAGACAGTACTAATCCTTCTGATTACACGTGGAACGATACTACGGGTAGCATCTCTGTAGGTGGTCGAAACTTGCTTGTAAAAACCAATCAAGGTATTACTAATTGGAATTGGCAGCTTTCCGATGGCGACAAGAGCGTTGAAGAAGTGAAAGTTGATGGCATTCGTGCTGTAAAACTAATCAAAGGTTCAACAGCAGCAAACACTGGGTGGAATTTCATTGAATATAATGGCTTGCTGCGTGAACTCATACAGCCGAAGTCGAAGTATGTTCTTTCGTTCGATGTTAAACCTAGCGTTGACGTAACTTTCTATGCAACGCTAGCACGAGGTGACTTTAACGAACCATTGACTGATACTGTCGCTATGCCTAAAGCATTAGCGAATCAGTGGAATAAGGTATCGTGCGTTTTGACAAGCAAAGAAACTTTGCCAAATATTGCATGGCAAGTTGTATACTTAGCAGGTATGCCAACAACAAACGGTAATTGGGTAATAATTAAAAATATCAAACTTGAAGAAGGTGACATACCTACTCAGTGGACACCTGCGATTGAGGACATACAAGATGAAATTGATTCCAAGGCCGATGCTGCTATGACGATTGAACAGATTAATGCACTTAATGAAAGGGCTGGGATCATTAAAGCAGAGATGGAAGCCAAAGCAAGCGCTGAAATTTTGAATAACTGGATTAAAAATTACCAAGATTTCGTTAAGGCAAACGAGACCGAGAGAGCTGCAGCCGAGAAAGCTTTGGTTAGCTCAAGTCAGCGGGTATCAACCATTGCTAAGGAATTAGGTGAACTGTCTGATCGTTGGAATTTCATCGATACTTACATGAACTCATCAAATGATGGGCTTGTGATTGGAAAGAATGACGGTAGCTCTAGCATGATGTTTAACCCTAACGGTCGCATTTCAATGTACTCGGCAGGGGAGGAGGTCATGTATATTTCGCAAGGTGTAATACACATCGAGAACGGGATCTTCTCGAAAACTATCCAAGTTGGTCGATATCGTGAGGAACAGTACCATCTTAACCCAGACATGAATGTCATTCGCTATGTAGGAGGTTTTTAA
Tertiary structure
PDB ID
f1b20c10530505ff9c988756f4fb9effdb6b1f516453c12d292b9ee8b9ef92fc
Model Confidence
Very high
pLDDT > 90
pLDDT > 90
High
90 > pLDDT > 70
90 > pLDDT > 70
Low
70 > pLDDT > 50
70 > pLDDT > 50
Very low
pLDDT < 50
pLDDT < 50