Protein
View in Explore- Genbank accession
- NP_059644.1 [GenBank]
- Protein name
- tail spike protein
- RBP type
-
TSPTFTSPTSPTSPTSPTSP
- Protein sequence
-
MTDITANVVVSNPRPIFTESRSFKAVANGKIYIGQIDTDPVNPANQIPVYIENEDGSHVQITQPLIINAAGKIVYNGQLVKIVTVQGHSMAIYDANGSQVDYIANVLKYDPDQYSIEADKKFKYSVKLSDYPTLQDAASAAVDGLLIDRDYNFYGGETVDFGGKVLTIECKAKFIGDGNLIFTKLGKGSRIAGVFMESTTTPWVIKPWTDDNQWLTDAAAVVATLKQSKTDGYQPTVSDYVKFPGIETLLPPNAKGQNITSTLEIRECIGVEVHRASGLMAGFLFRGCHFCKMVDANNPSGGKDGIITFENLSGDWGKGNYVIGGRTSYGSVSSAQFLRNNGGFERDGGVIGFTSYRAGESGVKTWQGTVGSTTSRNYNLQFRDSVVIYPVWDGFDLGADTDMNPELDRPGDYPITQYPLHQLPLNHLIDNLLVRGALGVGFGMDGKGMYVSNITVEDCAGSGAYLLTHESVFTNIAIIDTNTKDFQANQIYISGACRVNGLRLIGIRSTDGQGLTIDAPNSTVSGITGMVDPSRINVANLAEEGLGNIRANSFGYDSAAIKLRIHKLSKTLDSGALYSHINGGAGSGSAYTQLTAISGSTPDAVSLKVNHKDCRGAEIPFVPDIASDDFIKDSSCFLPYWENNSTSLKALVKKPNGELVRLTLATL
- Physico‐chemical
properties -
protein length: 667 AA molecular weight: 71855,98720 Da isoelectric point: 5,33857 aromaticity: 0,08846 hydropathy: -0,16432
Domains
Domains [InterPro]
IPR009093
ATT
1–113
ATT
1–113
IPR009093
ATT
1–100
ATT
1–100
IPR036730
ATT
2–110
ATT
2–110
G3DSA:2.170.14.10:FF:000001
Unmapped
2–110
Unmapped
2–110
IPR036730
ATT
7–109
ATT
7–109
1
667
Architecture
ATT 1-113 | STR 114-667
Legend:
ATT
STR
RBD
CBM
LEC
ENZ
CHP
LNK
TAS
TTP
UNK
Unmapped
Tail Spike Domain Segmentation
Tail Spike Domain Segmentation
This protein has been segmented into three structural domains: N-terminal, central domain, and C-terminal.
Domain Layout
1
667
| Domain | Start | End | Length (AA) | Confidence |
|---|---|---|---|---|
| N-terminal | 1 | 130 | 130 | 0,9594 |
| Central domain | 131 | 554 | 425 | 0,9864 |
| C-terminal | 555 | 667 | 112 | 0,8663 |
Legend:
N-terminal
Central domain
C-terminal
3D Structure with Domain Coloring
The structure is colored according to the domain segmentation: N-terminal (blue), Central (green), C-terminal (pink).
Domain Coloring
N-terminal
1-130
1-130
Central
131-554
131-554
C-terminal
555-667
555-667
Taxonomy
| Name | Taxonomy ID | Lineage | |
|---|---|---|---|
| Phage |
Salmonella phage P22 [NCBI] |
10754 | Uroviricota > Caudoviricetes > Lederbergvirus > |
| Host |
Salmonella enterica subsp. enterica serovar Typhimurium str. LT2 [NCBI] |
99287 | Bacteria > Proteobacteria > Gammaproteobacteria > Enterobacteriales > Enterobacteriaceae > Salmonella |
Coding sequence (CDS)
Coding sequence (CDS)
Genbank protein accession
NP_059644.1
[NCBI]
Genbank nucleotide accession
NC_002371.2
[NCBI]
CDS location
range 16204 -> 18207
strand +
strand +
CDS
ATGACAGACATCACTGCAAACGTAGTTGTTTCTAACCCTCGTCCAATCTTCACTGAATCCCGTTCGTTTAAAGCTGTTGCTAATGGGAAAATTTACATTGGTCAGATTGATACCGATCCGGTTAATCCTGCCAATCAGATACCCGTATACATTGAAAATGAGGATGGCTCTCACGTCCAGATTACTCAGCCGCTAATTATCAACGCAGCCGGTAAAATCGTATACAACGGCCAACTGGTGAAAATTGTCACCGTTCAGGGTCATAGCATGGCTATCTATGATGCCAATGGTTCTCAGGTTGACTATATTGCTAACGTATTGAAGTACGATCCAGATCAATATTCAATAGAAGCTGATAAAAAATTTAAGTATTCAGTAAAATTATCAGATTATCCAACATTGCAGGATGCAGCATCTGCTGCGGTTGATGGCCTTCTTATCGATCGAGATTATAATTTTTATGGTGGAGAGACAGTTGATTTTGGCGGAAAGGTTCTGACTATAGAATGTAAAGCTAAATTTATAGGAGATGGAAATCTTATTTTTACGAAATTAGGCAAAGGTTCCCGCATTGCCGGGGTTTTTATGGAAAGCACTACAACACCATGGGTTATCAAGCCTTGGACGGATGACAATCAGTGGCTAACGGATGCCGCAGCGGTCGTTGCCACTTTAAAACAATCTAAAACTGATGGGTATCAGCCAACCGTAAGCGATTACGTTAAATTCCCAGGAATAGAAACGTTACTCCCACCTAATGCAAAAGGGCAAAACATAACGTCTACGTTAGAAATTAGAGAATGTATAGGGGTCGAAGTTCATCGGGCTAGCGGTCTAATGGCTGGTTTTTTGTTTAGAGGGTGTCACTTCTGCAAGATGGTAGACGCCAATAATCCAAGCGGAGGTAAAGATGGCATTATAACCTTCGAAAACCTTAGCGGCGATTGGGGGAAGGGTAACTATGTCATTGGCGGACGAACCAGCTATGGGTCAGTAAGTAGCGCCCAGTTTTTACGTAATAATGGTGGCTTTGAACGTGATGGTGGAGTTATTGGGTTTACTTCATATCGCGCTGGGGAGAGTGGCGTTAAAACTTGGCAAGGTACTGTGGGCTCGACAACCTCTCGCAACTATAATCTGCAATTCCGCGACTCGGTCGTTATTTACCCCGTATGGGACGGATTCGATTTAGGTGCTGACACTGACATGAATCCGGAGTTGGACAGGCCAGGGGACTACCCTATAACCCAATACCCACTGCATCAGTTACCCCTAAATCACCTGATTGATAATCTTCTGGTTCGCGGGGCGTTAGGTGTAGGTTTTGGTATGGATGGTAAGGGCATGTATGTGTCTAATATTACCGTAGAAGATTGCGCTGGGTCTGGCGCGTACCTACTCACCCACGAATCAGTATTTACCAATATAGCCATAATTGACACCAATACTAAGGATTTCCAGGCGAATCAGATTTATATATCTGGGGCTTGCCGTGTGAACGGTTTACGTTTAATTGGGATCCGCTCAACCGATGGGCAGGGTCTAACCATAGACGCCCCTAACTCTACCGTAAGCGGTATAACCGGGATGGTAGACCCCTCTAGAATTAATGTTGCTAATTTGGCAGAAGAAGGGTTAGGTAATATCCGCGCTAATAGTTTCGGCTATGATAGCGCAGCGATTAAACTGCGGATTCATAAGTTATCAAAGACATTAGATAGCGGAGCATTGTACTCCCACATTAACGGGGGGGCCGGTTCTGGCTCAGCGTATACTCAACTTACTGCTATTTCAGGTAGCACACCTGACGCTGTATCATTAAAAGTTAACCACAAAGATTGCAGGGGGGCAGAGATACCATTTGTTCCTGACATCGCGTCAGATGATTTTATAAAGGATTCCTCATGTTTTTTGCCATATTGGGAAAATAATTCTACTTCTTTAAAGGCTTTAGTGAAAAAACCCAATGGAGAATTAGTTAGATTAACCTTGGCAACACTTTAG
Genome Context
Genome Context
Tertiary structure
1 / 30
PDB ID