Protein

Genbank accession
QHR65689.1 [GenBank]
Protein name
putative colanic acid-degrading protein
RBP type
TSP
Evidence DepoScope
Probability 1,00
TSP
Evidence RBPdetect2
Probability 0,95
Protein sequence
MATFPTYPATSLEQGVDLVIFSSNQLHDVINGSATESIETESGLIPTLRKALVDNFFFKSPLDWAQGTNETVFNQLRYFQNGVLSGYYYAPNATLVNPIPMQGTPVGDNNWVLWGLKTEQLATEVTPWVYSGATGYETVISPPYIFDSAIVTINGVIQVLGEAFRIEDSKIILSEPLGHDPATGLPNKLFAYIGKIVASADTDPNLENRLVSLETKTEELTESMDKVPLQTIARKYGLLDDEVAYATAGQSLTGIKALYDPVAQSSYTLPPNITGTLTSLSVSGVMTYSGGSVDLSALAVERGQFVHLPTSFGNNVTLTAKNQTIGRGAGRYRWAGSLPKSILASDTLETSGGLGPNAWVLTNTEGLLTPRGGTYNDLFDVVYLSEFANNTTTYTTPEAAFQAALLAAKASKSKILDAYGCDMTFTTSSFDIQGITLRGGVFRGQRDYRVQNATVEGTTFRNSRVMYWGGAVRMFDCLWDGAPRAGQVGSLVFQGNPISGTFEIDNCTFKNGLYGILQQGTGEPVTRGVFRNLTFMDMQGDAIELNVINKHYDDGCVIENIYLSNIDGTNAPIPLSNWGIGIGVAGKGPYGYGIPDDQYCKNITIRNVFAKRCRQIVHVEVGRNISIENIHGDPDQTVSVGTGLATGAVVMYGSKDFTIDGVYGEPKTDGSTLASNIRMIYLEWGTNAVLDEEGNPVVPAQGRPSNPCFNYSVRNIHTKTGRVFAGVSAGPGYENRVSFENIRCAALQLFGIASWLSMSNITCNIFDCVGQPESGPGTFYDGFFRREKSVLEMVNVNCYPDGKTMVTGRPQWSRCRYSDIHRVNCNVEANMYTNIAGGIGAIVGTTGKVYYLEPNPSRNIDGMHFPTGKEFDKGDMIVKADNTFFLVTTSGAYIPDIPAFGIRATQAGDTFLTQNLTPNGTATNASWLYHYPLSAGTRIRIPGAGVGGADLDTVITRAPYQDNDTWTNPVKIDIADPIVTPTSAGVRIKTIPNDSSNVSVGNTAVIRPTPVVDVPTT
Physico‐chemical
properties
protein length:1017 AA
molecular weight: 109867,27630 Da
isoelectric point:5,02504
aromaticity:0,09538
hydropathy:-0,13845

Domains

Domains [InterPro]
Legend: Pfam SMART CDD TIGRFAM HAMAP SUPFAM PRINTS Gene3D PANTHER Other

Taxonomy

  Name Taxonomy ID Lineage
Phage Escherichia phage nepoznato
[NCBI]
2696431 Viruses > Duplodnaviria > Heunggongvirae > Uroviricota > Caudoviricetes
Host No host information

Coding sequence (CDS)

Coding sequence (CDS)
Genbank protein accession
QHR65689.1 [NCBI]
Genbank nucleotide accession
MN850571 [NCBI]
CDS location
range 121979 -> 125032
strand -
CDS
ATGGCAACATTCCCAACATATCCAGCAACGTCTCTCGAACAGGGCGTAGATCTGGTTATTTTCTCGTCTAACCAACTCCATGATGTTATTAATGGAAGTGCTACAGAGAGTATTGAAACAGAATCAGGGTTGATACCAACCCTGAGAAAAGCCCTTGTTGACAACTTTTTCTTTAAAAGTCCATTAGATTGGGCGCAAGGCACAAACGAAACAGTTTTTAACCAACTTCGTTATTTCCAGAACGGAGTGTTGAGTGGTTATTACTATGCGCCTAATGCTACGCTGGTAAACCCAATTCCTATGCAAGGGACTCCAGTAGGAGATAACAACTGGGTATTGTGGGGACTGAAGACAGAACAACTAGCAACAGAAGTCACCCCTTGGGTATATTCTGGTGCAACTGGTTATGAAACAGTAATCAGTCCTCCCTACATCTTTGATAGTGCTATTGTTACAATCAATGGTGTTATTCAGGTTCTGGGAGAAGCTTTCCGCATAGAAGATAGTAAAATCATTTTATCTGAACCTTTGGGACACGATCCTGCAACAGGACTACCAAACAAGCTTTTTGCTTATATTGGTAAAATCGTAGCTTCAGCAGATACAGATCCTAATCTAGAGAACCGTCTTGTTAGTTTAGAAACAAAAACGGAAGAACTAACAGAAAGCATGGATAAAGTCCCTCTGCAAACTATTGCTCGTAAATATGGTTTGTTAGATGATGAGGTGGCTTACGCAACAGCAGGTCAATCGTTAACAGGGATTAAAGCTCTGTATGATCCAGTAGCACAGTCTTCCTACACATTACCTCCTAACATCACAGGAACTTTAACATCCTTATCTGTTTCTGGAGTAATGACTTATTCTGGGGGGTCCGTAGATCTTTCTGCACTGGCTGTTGAACGTGGACAGTTTGTTCATTTGCCTACAAGCTTTGGTAATAACGTAACTCTGACAGCTAAAAATCAGACGATTGGACGTGGTGCTGGTCGTTATCGTTGGGCTGGAAGTTTACCAAAAAGTATTTTGGCCTCTGATACATTGGAGACAAGCGGTGGATTAGGTCCAAATGCATGGGTTCTGACCAATACAGAAGGACTGTTAACCCCTCGTGGTGGAACATACAACGATTTATTCGATGTTGTCTACTTGTCCGAATTTGCTAACAATACAACTACCTATACCACACCAGAAGCAGCTTTCCAGGCGGCACTGTTAGCAGCAAAAGCTTCGAAAAGCAAGATTCTCGATGCTTACGGTTGCGATATGACTTTCACGACAAGCTCTTTTGATATTCAAGGTATCACTTTGCGTGGGGGTGTGTTCCGTGGTCAACGTGACTATCGTGTGCAAAATGCTACAGTAGAAGGCACAACTTTCCGTAACTCTCGTGTTATGTATTGGGGTGGTGCTGTGCGCATGTTCGATTGTTTGTGGGACGGTGCTCCTCGCGCAGGTCAAGTTGGTAGCTTGGTATTCCAAGGCAACCCAATATCAGGTACTTTTGAAATTGATAATTGTACTTTTAAAAACGGGTTATATGGCATCCTTCAACAAGGAACAGGGGAACCAGTAACACGCGGTGTCTTCCGTAACCTTACCTTCATGGACATGCAAGGTGATGCAATAGAACTGAACGTCATCAATAAACATTATGATGATGGTTGTGTGATTGAGAATATTTACCTGTCTAATATCGATGGGACTAACGCTCCTATTCCTCTGTCCAACTGGGGTATCGGTATTGGTGTAGCTGGTAAAGGACCATATGGTTACGGAATTCCTGATGATCAATATTGCAAGAATATTACAATTCGTAACGTTTTTGCAAAACGTTGTCGTCAGATTGTCCACGTAGAAGTTGGTCGTAATATCTCTATCGAAAACATTCATGGCGATCCAGACCAAACTGTATCTGTTGGGACAGGTTTGGCAACTGGCGCAGTTGTAATGTATGGAAGTAAAGACTTTACGATTGATGGTGTTTATGGAGAGCCTAAAACAGACGGTAGCACACTTGCAAGTAATATCCGTATGATTTATTTGGAGTGGGGAACTAACGCAGTTCTGGACGAAGAGGGTAATCCGGTAGTTCCGGCACAGGGTAGACCTTCAAATCCATGCTTTAACTATTCTGTTCGTAACATTCACACCAAAACTGGGCGTGTTTTTGCGGGAGTTTCCGCTGGGCCAGGTTATGAAAACAGAGTTAGCTTTGAAAATATTCGTTGCGCCGCATTGCAGTTGTTCGGTATAGCTTCTTGGCTTTCAATGTCTAACATCACCTGTAATATCTTCGATTGTGTTGGTCAACCTGAATCTGGTCCAGGGACATTTTATGATGGATTCTTCCGTAGAGAAAAATCTGTTCTTGAGATGGTTAACGTAAACTGTTACCCAGATGGCAAGACGATGGTTACTGGTCGTCCACAGTGGAGCCGTTGCCGTTACTCGGATATTCATCGTGTCAACTGTAATGTTGAAGCTAATATGTATACCAATATCGCAGGTGGTATTGGAGCTATTGTAGGGACAACTGGCAAAGTTTATTACTTAGAGCCTAACCCTAGCCGAAACATTGATGGGATGCATTTCCCAACAGGGAAGGAGTTTGACAAAGGGGATATGATAGTCAAAGCGGATAATACGTTCTTCCTTGTAACAACAAGTGGGGCATACATTCCAGATATCCCAGCTTTCGGTATTCGAGCTACACAAGCAGGAGATACCTTCCTGACCCAGAACTTGACTCCTAATGGGACAGCAACAAACGCTTCTTGGTTGTATCACTATCCTCTGTCAGCAGGAACACGTATCCGTATTCCTGGTGCTGGTGTAGGTGGAGCAGATTTGGACACAGTTATCACTCGTGCACCTTATCAAGATAATGATACGTGGACTAACCCAGTTAAAATAGACATTGCAGATCCAATTGTAACGCCTACGAGTGCTGGAGTTAGAATCAAAACCATACCTAACGATTCTAGTAATGTTTCTGTAGGTAATACTGCGGTTATTCGCCCAACACCAGTGGTAGATGTGCCAACAACCTAA

Gene Ontology

No Gene Ontology terms available.

Enzymatic activity

No enzymatic activity data available.

Tertiary structure

PDB ID
d1bfb5d1661ae79e66b98bb5c573c8249be675beec531677a2953275494625a6
ESMFold
Source ESMFold
Method ESMFold
Resolution 0,6635
Evidence 0,6635

Literature

Title Authors Date PMID Source
Exploring the Remarkable Diversity of Culturable Escherichia coli Phages in the Danish Wastewater Environment Olsen,N.S., Forero-Junco,L., Kot,W. and Hansen,L.H. 2020 GenBank