Protein
View in Explore- Genbank accession
- XYY30176.1 [GenBank]
- Protein name
- tail fiber protein
- RBP type
-
TFTSPTF
- Protein sequence
-
MQSLLETTSKGDRNPSSVRLLIQLQRNGNWVTEKDVTINGKTTSQFLASVILDNLPPRPFNIRMVRETADSTSDQLQNKTLWSSYTEIIDVKQCYPNTAIVGLQVDAEQFGGQQMTVNYHIRGRIIQVPSNYDPEKRTYSGIWDGSLKPAYSNNPAWCLWDMLTHPRYGMGKRLGAADVDKWALYAIAQYCDQMVPDGFGGTEPRMTFNAYLSQQRKAWDVLSDFCSAMRCMPVWNGQTLTFVQDSPSDVVWPYTNSDVVVDDNGVGFRYSFSALKDRHTAVEVNYTDPQNGWQTSTELVEDPEAILRYGRNLLKMDAFGCTSRGQAHRAGLWVIKTGLLETQTVDFTLGSQGLRHTPGDIIEICDNDYAGTMTGGRILSIDAASRTLTLDREVTLPETGAATVNLINGSGKPVSVAITAHPAPDRIQVSTLPDGVETYGVWGLSLPSLRRRLFRCVSIRENTDGTFAITAVQHVPEKEAIVDNGASFEPQSGTLNSVIPPAVQHLTVEVSAADGQYLAQAKWDTPKVVKGVSFMLRLTVAADDGSERLVSTARTTETTYRFTQLAPGNYRLTVRAVNAWGQQGDPASVSFRIAAPAAPSQIELTPGYFQITAVPKLAVYDPTVQFEFWFSEAKIADTSQVETSARYLGTGSQWSVSGPHIKPGKDFWFYVRSVNLVGKSAFVEVSGQPSNDGEGYLEFFREKIGKLHLAQGLWELIDNSQLADEMAEMKTSITETRNEITQTVSKTLEDQSATIQQIQRVQKDTNDDLAALYMLKVQKTKDGIPYVAGIGAGIEDTDGQPLSNILLLADRIAMINPESGNSTPLFVAQGNQLFMNDVFLKRLFAVSITSSGNPPAFSLTPDGRLTAKNADISGSVNANSGTLNNVTINENCQIKGKLSANQIEGDIVKTVSKSFPRTSTYASGTITVRISDDQKFDRQVMIPPVLFRGGKHENFNSNNQQSYWYSTCRLRVTRNGQEIFNQSTTDAQGVFSSVIDMPAGQGTLTLTFTVSSSGANNWTPTTSISDLLVVVMKKSTAGISIS
- Physico‐chemical
properties -
protein length: 1042 AA molecular weight: 114376,22500 Da isoelectric point: 5,49709 aromaticity: 0,08349 hydropathy: -0,29386
Domains
Domains [InterPro]
DC_0248
STR
1–1037
STR
1–1037
IPR055385
ATT
2–90
ATT
2–90
IPR053171
Unmapped
2–728
Unmapped
2–728
IPR055383
STR
492–596
STR
492–596
IPR036116
STR
499–601
STR
499–601
IPR003961
STR
500–592
STR
500–592
IPR003961
STR
502–597
STR
502–597
1
1042
Architecture
ATT 1-90 | STR 91-212 | ATT 213-381 | STR 382-597 | ATT 598-700 | STR 701-1042
Legend:
ATT
STR
RBD
CBM
LEC
ENZ
CHP
LNK
TAS
TTP
UNK
Unmapped
Tail Spike Domain Segmentation
Tail Spike Domain Segmentation
This protein has been segmented into three structural domains: N-terminal, central domain, and C-terminal.
Domain Layout
1
1042
| Domain | Start | End | Length (AA) | Confidence |
|---|---|---|---|---|
| N-terminal | 1 | 899 | 899 | 0,9467 |
| Central domain | 900 | 1031 | 133 | 0,1975 |
| C-terminal | 1032 | 1042 | 10 | 0,9835 |
Note: Constraints were applied during segmentation.
Fixed 55 C-terminal predictions appearing before Central domain|C-terminal too short, adjusted boundary
Fixed 55 C-terminal predictions appearing before Central domain|C-terminal too short, adjusted boundary
Legend:
N-terminal
Central domain
C-terminal
3D Structure with Domain Coloring
The structure is colored according to the domain segmentation: N-terminal (blue), Central (green), C-terminal (pink).
Domain Coloring
N-terminal
1-899
1-899
Central
900-1031
900-1031
C-terminal
1032-1042
1032-1042
Taxonomy
| Name | Taxonomy ID | Lineage | |
|---|---|---|---|
| Phage |
Escherichia phage vB_EcoS_phi330 [NCBI] |
3411874 | Viruses > Duplodnaviria > Heunggongvirae > Uroviricota > Caudoviricetes |
| Host |
Escherichia coli [NCBI] |
562 | cellular organisms > Bacteria > Pseudomonadati > Pseudomonadota > Gammaproteobacteria > Enterobacterales |
Coding sequence (CDS)
Coding sequence (CDS)
Genbank protein accession
XYY30176.1
[NCBI]
Genbank nucleotide accession
PV189401.1
[NCBI]
CDS location
range 17482 -> 20610
strand +
strand +
CDS
GTGCAGTCACTGTTGGAGACCACCTCAAAGGGCGACCGTAATCCCTCTTCTGTCCGACTGCTGATTCAGTTGCAGCGTAACGGTAACTGGGTGACGGAAAAGGATGTCACCATTAACGGCAAGACCACCTCGCAGTTTCTGGCGTCGGTGATTCTGGATAATCTGCCTCCCCGTCCTTTTAACATCCGGATGGTCCGGGAGACGGCGGACAGCACCTCGGACCAGCTGCAGAATAAGACGCTCTGGTCGTCATACACCGAAATCATCGATGTGAAACAGTGCTACCCGAACACGGCGATTGTGGGGCTGCAGGTGGATGCGGAGCAGTTTGGTGGCCAGCAGATGACGGTGAACTACCATATCCGAGGTCGCATCATCCAGGTGCCGTCAAACTATGACCCGGAAAAACGCACGTACAGCGGCATCTGGGACGGCAGCCTGAAACCGGCATACAGCAACAACCCGGCCTGGTGCCTGTGGGACATGCTGACTCACCCGCGCTACGGCATGGGAAAACGCCTGGGGGCGGCGGATGTGGACAAGTGGGCGCTGTATGCCATTGCGCAGTACTGCGACCAGATGGTCCCTGATGGTTTCGGGGGCACAGAGCCGCGGATGACCTTTAATGCGTACCTGTCACAACAGCGTAAGGCGTGGGACGTTCTCAGTGATTTCTGCTCGGCGATGCGCTGTATGCCGGTATGGAACGGCCAGACGCTGACGTTCGTTCAGGACAGCCCGTCGGATGTGGTGTGGCCGTACACCAACAGTGATGTGGTGGTGGATGATAACGGCGTGGGGTTTCGCTACAGCTTCAGCGCCCTGAAGGACCGCCACACGGCGGTGGAGGTGAATTACACCGACCCGCAGAACGGCTGGCAGACCTCCACGGAACTGGTGGAAGACCCGGAAGCCATACTGCGCTACGGGCGCAACCTGCTGAAGATGGATGCGTTCGGTTGCACCAGTCGCGGTCAGGCCCACCGTGCCGGGCTGTGGGTGATAAAGACCGGACTGCTGGAAACGCAGACGGTGGATTTCACGCTCGGGTCACAGGGGCTGCGTCACACACCCGGTGACATTATTGAAATCTGTGATAACGACTATGCCGGGACCATGACCGGCGGACGTATCCTGTCCATCGATGCCGCCAGCCGCACCCTGACACTGGACCGTGAGGTGACCCTGCCGGAGACCGGTGCCGCCACGGTGAACCTGATTAACGGCAGCGGTAAGCCGGTGAGCGTGGCCATCACTGCACACCCCGCGCCGGACCGGATACAGGTCAGCACCCTGCCTGATGGTGTGGAGACATACGGTGTATGGGGACTCTCCCTGCCGTCACTGCGTCGTCGCCTGTTCCGCTGTGTCTCCATCCGGGAAAACACGGACGGCACCTTTGCCATCACGGCGGTGCAGCACGTACCGGAAAAAGAAGCCATCGTGGATAACGGGGCCAGCTTTGAGCCGCAGTCAGGCACCCTGAACAGCGTTATTCCACCGGCAGTGCAGCACCTGACGGTGGAGGTGAGCGCGGCTGACGGTCAGTATCTGGCACAGGCGAAATGGGACACGCCGAAGGTGGTGAAGGGCGTGAGCTTTATGCTTCGCCTGACCGTGGCCGCGGATGACGGCAGTGAGCGGCTGGTCAGCACGGCCCGGACGACGGAAACCACATACCGCTTCACGCAACTGGCGCCGGGGAACTACAGGCTGACAGTCCGGGCGGTAAATGCGTGGGGGCAGCAGGGCGATCCGGCGTCGGTATCGTTCCGGATTGCCGCACCGGCAGCGCCGTCACAGATTGAGCTGACACCGGGCTATTTTCAGATAACGGCGGTCCCGAAACTGGCTGTATATGACCCGACGGTACAGTTTGAATTCTGGTTTTCGGAGGCAAAAATCGCAGACACATCTCAGGTGGAAACCTCTGCCCGTTATCTGGGGACCGGCAGTCAGTGGAGTGTATCCGGCCCGCACATTAAGCCCGGGAAGGATTTCTGGTTTTACGTGCGCAGCGTCAACCTGGTGGGGAAATCTGCGTTTGTGGAAGTCAGCGGGCAGCCCAGCAATGATGGTGAAGGGTATCTGGAATTTTTCCGGGAAAAAATAGGAAAACTGCATCTGGCTCAGGGGTTGTGGGAACTGATAGATAACAGCCAGCTTGCGGATGAGATGGCGGAGATGAAGACCAGCATCACGGAAACCCGCAATGAAATCACACAGACGGTCAGTAAAACGCTGGAGGACCAGAGCGCCACCATACAGCAGATACAGCGCGTGCAGAAGGACACAAATGATGACCTGGCTGCGCTGTACATGCTGAAGGTTCAAAAAACGAAAGACGGCATTCCCTATGTGGCCGGGATTGGTGCAGGGATTGAGGATACTGATGGCCAGCCACTGAGCAACATACTGCTGCTGGCTGACCGTATCGCGATGATAAATCCGGAGAGCGGCAACAGCACGCCGTTATTTGTGGCGCAGGGGAATCAGCTGTTCATGAACGACGTGTTCCTGAAACGACTGTTTGCGGTGAGCATCACATCATCCGGCAATCCTCCGGCATTTTCCCTGACGCCGGACGGGCGACTGACGGCGAAAAATGCGGATATCAGTGGCAGTGTGAATGCGAACTCAGGGACGCTCAACAACGTCACGATTAATGAGAACTGTCAGATTAAGGGGAAACTGTCAGCCAACCAGATTGAAGGCGATATTGTCAAAACGGTCAGCAAGTCTTTCCCCCGCACGAGCACTTATGCCAGTGGCACCATCACGGTAAGAATCAGTGATGATCAGAAGTTTGACCGGCAGGTCATGATACCGCCAGTGTTATTCCGCGGTGGTAAGCATGAGAATTTCAACAGTAATAACCAACAGTCATACTGGTATTCAACCTGCCGGTTAAGAGTGACCCGCAATGGTCAGGAGATTTTTAATCAGTCCACGACGGATGCTCAGGGCGTATTTTCTTCAGTTATAGATATGCCTGCCGGACAGGGGACGCTGACACTGACATTCACCGTATCTTCATCAGGAGCGAATAACTGGACACCAACAACCAGTATCAGCGATCTGCTGGTTGTGGTGATGAAAAAATCCACAGCAGGTATCAGTATCAGCTGA
Genome Context
Genome Context
Tertiary structure
PDB ID
1f1fd33400d2d855b5528fa1005f747daad2614505c7c8d7f60670740866e287
Model Confidence
Very high
pLDDT > 90
pLDDT > 90
High
90 > pLDDT > 70
90 > pLDDT > 70
Low
70 > pLDDT > 50
70 > pLDDT > 50
Very low
pLDDT < 50
pLDDT < 50