Protein

Genbank accession
QFR58941.1 [GenBank]
Protein name
tail fiber protein
RBP type
TSP
Evidence RBPdetect
Probability 0,89
Protein sequence
MALIYHYTRNEDGTFNVVRYRDNPMNFVVNNVPDGVPIRIFIDEICEDNDVTEDFDALKEDAIFYIVESAGGGVVKGVMKIFSVILKPLAKLLSPSAKGLSSNFANSQADSPNNSLTDRNNKPRPYERSYDICGTVQTVPNNLMTTYKIFNDTGRIIEYGYYDAGRGHLDIKAEDVTDGDTRVQDIDGSSVAIYAPYTSPNNTSTPQLQIGEVIDQQLYTTIESNEIDGITLRAPNDLGVTFSGETAVVTLSGTTGRITETAGGVDFTEDFSVGDIANLNAWTTNTVNIGDHYTVTSVSETVIELDVSSRLLSWSSASGNPIRGSDGENSIRPTDTLDKTLTDWVSINRSTVERIVVNIAAPNGMYKDNGSKLTMSATAEVQYQALDESGAPFGPIYTATGTITGRSSDYNGVTIYGYLPTASRVRVRARRTSDTDRGFNGTVSDEIVYANLYGQSVDTTPHYGNRTTVHSARKQTPRATEVRQPQLRMIATEMVFKYLGGGVFDTVMTPNRQAVQSLIRLARDPVVGGLDLSVANMDALLETQEEVESYFGSSDAGEFCYTFDDYKTTMQEIVTIIAEAIFCTPYRKGADILLDFERPRLGPEMVFTHRSKVKTSEKWSRTFNDPQVFDSLKFSYIDPVTNVKETISIPETGGLKTETYDSKGVRNKKQAYWLAHRRHQKNILRKVNVSFTATEEGIFARPNRPISVVKGSRMATYDGYITSVDGLTVHLSSPVQFTAGEPHSLVLKTRDGGAQSIPVIEGPDNRSVIMLSVPQEAIYTGNSALKTEFSFGSDSRHAAQMIMVSTVEPSDDRTVRITGFNYDADYYKYDGVSPFGRAFSDGFSNGFS
Physico‐chemical
properties
protein length:848 AA
molecular weight: 93374,84020 Da
isoelectric point:5,00032
aromaticity:0,09080
hydropathy:-0,36922

Domains

Domains [InterPro]
QFR58941.1
1 848
Legend: Pfam SMART CDD TIGRFAM HAMAP SUPFAM PRINTS Gene3D PANTHER Other

Taxonomy

  Name Taxonomy ID Lineage
Phage Salmonella virus VSiA
[NCBI]
2653661 Viruses > Duplodnaviria > Heunggongvirae > Uroviricota > Caudoviricetes
Host No host information

Coding sequence (CDS)

Coding sequence (CDS)
Genbank protein accession
QFR58941.1 [NCBI]
Genbank nucleotide accession
MN393079.1 [NCBI]
CDS location
range 19262 -> 21808
strand +
CDS
TTGGCGCTGATATACCACTACACAAGAAACGAGGATGGTACTTTTAACGTGGTTCGATACCGCGACAACCCTATGAATTTTGTAGTAAATAATGTGCCTGATGGGGTTCCTATTCGCATATTTATCGATGAAATATGCGAAGATAACGACGTAACAGAAGATTTTGATGCGTTAAAAGAAGATGCAATCTTTTATATTGTTGAGTCTGCGGGGGGTGGCGTTGTAAAAGGGGTAATGAAAATATTCAGCGTCATCCTTAAACCACTTGCAAAGCTTTTAAGCCCATCGGCAAAAGGCTTATCTTCTAATTTTGCAAACTCGCAGGCGGACTCGCCAAACAATAGCCTTACAGACCGCAACAATAAGCCACGACCATATGAGCGCAGTTATGACATATGTGGCACGGTACAAACCGTACCGAACAATCTTATGACAACGTACAAAATATTCAATGATACGGGCCGCATCATTGAATATGGGTATTATGACGCCGGGAGAGGTCATCTCGATATAAAAGCTGAAGATGTTACTGATGGGGACACGCGGGTGCAGGATATTGACGGGTCGTCCGTCGCGATATACGCCCCGTACACATCACCTAATAACACATCAACACCACAGTTACAGATAGGAGAGGTTATAGACCAGCAGTTGTACACCACTATAGAGTCCAACGAGATAGACGGCATAACGCTTAGGGCCCCAAATGATTTAGGCGTAACATTCAGTGGAGAGACGGCAGTAGTTACGTTGTCAGGAACAACAGGAAGAATAACAGAAACTGCCGGTGGTGTTGACTTCACAGAAGATTTTTCTGTAGGCGATATCGCTAACTTAAATGCCTGGACAACTAACACCGTTAATATAGGAGACCATTACACTGTAACATCGGTGTCAGAAACAGTTATTGAACTGGACGTATCATCGAGGCTACTGTCTTGGTCAAGCGCAAGCGGCAACCCTATCAGGGGGAGCGACGGTGAAAATAGTATCAGACCAACGGACACGCTAGATAAAACATTAACGGACTGGGTGTCTATAAATAGGAGTACCGTAGAGCGCATAGTGGTTAATATCGCCGCGCCAAATGGTATGTATAAAGACAATGGCAGCAAATTAACGATGTCTGCGACCGCCGAGGTGCAATATCAGGCCCTGGATGAGAGTGGAGCGCCGTTCGGGCCCATATACACCGCCACGGGAACAATAACAGGTCGTAGCTCAGATTATAATGGAGTTACGATATATGGCTATCTACCGACCGCGTCCCGTGTCCGGGTTAGAGCGCGGCGCACATCAGACACCGATAGAGGTTTTAATGGCACCGTTTCGGACGAAATTGTATACGCTAACCTGTACGGTCAGTCTGTGGACACAACCCCACACTATGGTAATAGAACAACAGTACATTCGGCGCGAAAACAAACCCCCAGGGCGACGGAAGTTAGGCAACCACAGCTACGTATGATTGCCACCGAAATGGTATTCAAATACTTGGGTGGTGGAGTTTTTGATACGGTAATGACACCAAATAGACAAGCTGTTCAGTCTCTTATTCGGCTAGCGAGGGACCCGGTAGTTGGTGGTCTTGATTTATCTGTAGCTAATATGGACGCGCTTTTAGAGACTCAAGAAGAAGTCGAGTCGTATTTCGGTTCTTCTGATGCTGGCGAGTTTTGCTACACGTTCGATGATTATAAGACCACAATGCAGGAAATAGTTACCATTATTGCCGAGGCTATATTTTGCACTCCGTATCGAAAAGGTGCGGATATACTTTTAGATTTTGAACGTCCGCGATTAGGGCCGGAGATGGTGTTTACTCACCGCAGTAAAGTAAAAACGTCTGAGAAATGGTCACGAACATTCAATGACCCCCAGGTTTTTGACAGCCTTAAATTCTCGTACATAGACCCCGTAACAAACGTTAAGGAGACAATATCTATCCCTGAAACAGGAGGTTTAAAGACCGAGACGTACGACTCTAAAGGGGTTCGCAATAAAAAACAAGCATACTGGCTGGCGCACAGACGGCATCAAAAAAATATCTTACGCAAAGTTAACGTATCATTCACCGCAACCGAAGAAGGTATTTTTGCACGACCAAACCGCCCGATATCTGTGGTTAAAGGCTCGCGCATGGCTACTTATGATGGCTATATAACTTCAGTAGACGGTTTAACGGTTCACTTGTCGTCGCCAGTACAGTTTACCGCAGGGGAACCACATTCGCTGGTACTGAAAACTCGCGATGGTGGCGCGCAAAGCATTCCGGTAATAGAAGGGCCTGACAATCGTTCTGTAATAATGCTTTCAGTACCGCAAGAAGCCATCTATACTGGTAACAGCGCGCTGAAGACAGAATTTTCCTTCGGCAGCGATAGCCGCCATGCGGCGCAAATGATTATGGTCAGCACCGTAGAACCGAGCGACGATAGAACCGTGCGCATCACTGGGTTCAATTACGATGCCGATTACTACAAATACGATGGCGTTTCGCCATTTGGCCGCGCATTTAGCGACGGATTCAGCAACGGTTTTAGTTAA

Gene Ontology

No Gene Ontology terms available.

Enzymatic activity

No enzymatic activity data available.

Tertiary structure

PDB ID
ec45c440ec8e5ab3ea33ae14d845dda84904feee133b54b4a2bee9d6ab39751c
ESMFold
Source ESMFold
Method ESMFold
Resolution 0,8479
Evidence 0,8479

Literature

Title Authors Date PMID Source
Complete genome sequence of Salmonella Infantis bacteriophage VSiA Denisenko,E., Kislichkina,A., Verevkin,V., Krasilnikova,V. and Volozhantsev,N. 2018-05-24 GenBank