Genbank accession
WPJ72496.1 [GenBank]
Protein name
L-shaped tail fiber protein
RBP type
TSP
Evidence DepoScope
Probability 1,00
TF
Evidence RBPdetect
Probability 0,88
TF
Evidence RBPdetect2
Probability 0,95
TF
Evidence GenBank
Probability 1,00
TF
Evidence Phold
Probability 1,00
Protein sequence
MALKTKIIVQQILNIDDTTTTASKYPKYTVVLGNSISSITAGELTAAVEAAAESAAAAKDSEIAAKDSENKAKDSEIQAGIHAGASEASATQSAASAAESERQANLSQGSAENSAASALESKNFKDASELAAQNAEQSKILAEQAQRAAEAAQSGAKASENKASAFATQAAASSASAGDFAAAAKQSELNAKTSETNAATSEVEAETQAETATTEANRAKAEADRAAQIVDSKLDKEDISGFIKVYKTKEEADADVSSRVLGEKILVWNQTDSKYGWYKVAGTAEAPVLELVETEQKLVSINNVRADDAGNVQITLPGGNPSLWLGEVTWFPYDKDSGVGYPGVLPADGREVLRVDYPDTWEAIEAGLIPSVTEEQWQAGATLYFSTGNGTTTFRLPDMMQGQAFRAAAKGEENAGNIKEQIPYITMINGKAPADDGTITLGNAADKNVWNGIDGEVLLRGAFGLGGTGLILNEPDAVSFFKAMRAFGSGYYRNDSESNPVIPKYSAGFYSKTADTHTFICSAYGNGVTFAATINDALLDGENPTVHTNILYGTANKPDLNTDTQGVLGVEKGGTGATTQKGARLNLDTPVGSRAIGMPNNSDVLAFMKSSAESGYYSSGNIVTGVPETAGWYMFDLHVHGKNAAGEMEYGNVYCTTSAGAIWYTLMEVGVWQPWRRLTTEHGIIPITSGGTGTNNANDARINLGLGPINAPTFSGMTLQGTNETTSGIAVFSNRNAEGTQLSYSRMYHEIQSGVGKTTIQTTREGGATNYFQIDEYGNIGNINSIIAYGYMGLGAANAMGNASIAIGDSDSGLKWNSDGNISTVADGVKIATWTPHGFYTHKIISSDVANTERGMYVNGVRTTGASALVAGVIEAGSHVGWRDRASGMLVELNTRGAAANIWKATRWGDQHAGASDIVIYDDGSPYYRTLVGGGEFGFNGLGQATCTSWISTSDIRLKAQLKEIVSAKDKVKSLQGYTYFKRNSLVEDEHSFYCEEAGLIAQDVQTVLPEAVYKIANSDLLGVNYSGVTALLANAVKEMLADAEAQEARISNLEEELAELKALIATLVNK
Physico‐chemical
properties
protein length:1071 AA
molecular weight: 113026,69140 Da
isoelectric point:4,76910
aromaticity:0,07470
hydropathy:-0,27890

Domains

Domains [InterPro]
WPJ72496.1
1 1071
Legend: Pfam SMART CDD TIGRFAM HAMAP SUPFAM PRINTS Gene3D PANTHER Other

Taxonomy

  Name Taxonomy ID Lineage
Phage Salmonella phage CRW-SP6
[NCBI]
3079602 Viruses > Duplodnaviria > Heunggongvirae > Uroviricota > Caudoviricetes
Host No host information

Coding sequence (CDS)

Coding sequence (CDS)
Genbank protein accession
WPJ72496.1 [NCBI]
Genbank nucleotide accession
OR464144.1 [NCBI]
CDS location
range 20410 -> 23625
strand +
CDS
ATGGCACTTAAAACTAAAATTATTGTACAGCAGATTCTGAACATAGATGACACTACAACTACTGCTAGTAAGTATCCTAAATATACAGTAGTTTTAGGTAATTCTATTAGTTCTATTACTGCTGGTGAACTAACAGCGGCTGTTGAAGCCGCCGCAGAGTCTGCTGCTGCTGCTAAAGATTCTGAAATAGCAGCTAAAGACTCTGAAAATAAAGCTAAAGATTCGGAAATTCAAGCGGGTATTCATGCTGGTGCTTCTGAGGCTTCAGCAACCCAGTCTGCTGCTTCTGCTGCTGAATCTGAAAGACAAGCTAACTTATCTCAAGGTAGTGCGGAAAACTCTGCTGCTTCTGCTTTAGAATCTAAGAATTTTAAAGATGCTTCGGAACTTGCTGCTCAAAATGCAGAGCAGAGTAAGATTTTAGCAGAGCAAGCTCAAAGAGCGGCAGAAGCTGCCCAGTCTGGTGCTAAAGCTTCTGAAAATAAAGCATCAGCATTTGCTACACAAGCTGCTGCATCTTCAGCTTCCGCAGGAGATTTTGCTGCAGCCGCTAAACAATCTGAATTAAATGCTAAAACTTCTGAAACCAATGCCGCAACATCAGAAGTGGAAGCGGAAACCCAAGCTGAAACTGCTACTACTGAGGCAAATCGTGCTAAGGCTGAAGCCGATCGCGCAGCTCAGATTGTAGATAGTAAGTTAGATAAAGAAGATATATCTGGCTTTATCAAAGTCTACAAGACTAAAGAAGAAGCGGACGCCGACGTTAGTAGCCGCGTACTAGGTGAAAAGATCCTAGTGTGGAACCAAACTGACTCAAAATATGGATGGTATAAAGTAGCTGGAACTGCTGAGGCTCCAGTATTAGAGTTAGTAGAGACAGAGCAAAAGCTAGTTTCTATTAATAACGTTCGTGCAGATGACGCAGGTAACGTACAGATTACTCTTCCTGGTGGTAATCCTTCCTTATGGTTGGGTGAAGTTACTTGGTTCCCTTATGACAAAGATTCAGGTGTTGGCTATCCTGGTGTTCTCCCTGCTGATGGCCGCGAAGTCCTTCGTGTAGACTATCCAGATACGTGGGAGGCTATCGAAGCCGGTCTGATTCCTTCTGTTACTGAAGAACAATGGCAAGCTGGTGCAACTCTCTACTTCTCCACTGGTAATGGTACTACTACTTTCCGCCTACCTGATATGATGCAGGGCCAAGCATTCCGTGCTGCTGCAAAAGGAGAGGAAAACGCTGGTAATATTAAAGAGCAAATCCCGTACATCACTATGATTAATGGTAAAGCTCCTGCTGACGATGGTACAATTACTTTAGGTAATGCTGCGGATAAAAACGTATGGAATGGTATTGATGGTGAAGTACTGTTAAGAGGTGCTTTTGGTCTTGGAGGTACTGGTTTAATTCTTAATGAACCTGATGCTGTTTCCTTCTTTAAAGCAATGCGTGCTTTTGGTTCAGGATATTATAGAAATGACTCTGAAAGTAACCCAGTAATCCCTAAATACTCTGCAGGATTCTACTCCAAAACTGCCGACACTCATACTTTTATCTGTTCTGCTTATGGTAATGGTGTTACTTTCGCAGCTACTATAAATGATGCATTATTAGATGGAGAAAATCCTACTGTACATACAAATATTCTTTATGGTACAGCAAATAAACCTGATCTGAATACCGATACTCAAGGAGTTTTAGGAGTAGAGAAGGGCGGTACTGGTGCTACTACGCAGAAAGGTGCTAGACTAAATCTGGATACTCCTGTAGGCAGCAGAGCTATTGGAATGCCTAATAACTCTGATGTACTAGCTTTCATGAAATCTTCCGCAGAAAGCGGATATTATTCCTCTGGTAATATAGTTACTGGAGTTCCAGAAACTGCAGGATGGTATATGTTCGATCTCCATGTACATGGTAAGAATGCTGCGGGAGAAATGGAGTATGGTAATGTATACTGTACAACAAGTGCTGGTGCTATTTGGTACACCTTAATGGAGGTTGGTGTATGGCAGCCATGGAGACGTTTGACCACAGAACATGGTATTATTCCTATTACTTCAGGGGGTACTGGTACAAATAATGCAAATGACGCAAGAATAAATCTAGGTCTTGGTCCTATAAATGCACCTACTTTTAGTGGTATGACTCTTCAGGGTACTAATGAAACTACTTCAGGTATAGCGGTTTTTAGTAATAGAAATGCGGAAGGGACTCAACTTTCCTATTCTAGAATGTACCATGAAATTCAGAGTGGTGTTGGTAAAACTACTATTCAGACTACAAGAGAGGGCGGGGCGACTAACTATTTCCAAATTGATGAGTATGGTAATATTGGGAATATTAACTCAATTATTGCATATGGATACATGGGATTAGGTGCTGCTAATGCTATGGGAAACGCCTCTATTGCGATTGGTGACTCTGACTCTGGGCTAAAATGGAATAGTGATGGTAACATAAGTACTGTAGCAGATGGTGTAAAAATAGCCACATGGACACCTCATGGATTTTATACACATAAAATAATAAGCTCAGATGTTGCTAATACCGAAAGAGGGATGTATGTAAACGGGGTTAGGACTACCGGTGCCTCCGCTCTTGTAGCTGGGGTTATAGAAGCTGGATCTCATGTTGGTTGGAGAGATAGAGCTTCAGGTATGCTTGTTGAATTGAATACTAGAGGAGCTGCTGCCAATATCTGGAAAGCAACTAGATGGGGTGACCAACATGCTGGTGCATCTGACATCGTTATTTATGATGATGGATCTCCTTATTATAGAACTCTTGTAGGCGGTGGTGAATTTGGGTTCAATGGCCTTGGACAAGCTACCTGTACTTCTTGGATCAGTACATCTGATATTAGGCTTAAGGCACAGCTAAAAGAGATAGTATCTGCTAAAGATAAGGTAAAATCCCTACAGGGGTACACTTATTTTAAACGTAATAGTTTGGTTGAAGATGAGCATTCCTTTTATTGTGAAGAGGCAGGATTAATCGCACAAGATGTTCAAACTGTACTACCTGAAGCTGTATATAAAATAGCTAACTCAGATCTTCTCGGTGTTAATTACTCTGGTGTTACCGCATTATTGGCTAACGCAGTAAAAGAGATGTTGGCGGATGCGGAGGCTCAGGAAGCTCGTATCAGTAATCTAGAAGAAGAACTGGCAGAGTTAAAAGCTCTAATAGCCACTCTGGTAAATAAGTAA

Tertiary structure

PDB ID
cbd81e91f2f4e52201bb37b9534f0308512ccb2dc09c320df1faf69b73e3f712
ESMFold
Source ESMFold
Method ESMFold
Resolution 0,6074
Oligomeric State monomer
Model Confidence
Very high
pLDDT > 90
High
90 > pLDDT > 70
Low
70 > pLDDT > 50
Very low
pLDDT < 50