Protein

Genbank accession
QAY00559.1 [GenBank]
Protein name
hypothetical protein
RBP type
TSP
Evidence DepoScope
Probability 1,00
Protein sequence
MSKKIYGGLDVTHNVTVKGKRTVQGINDQVFDANGNMEIDTYSKEELLDVISLLPISHYGTHNYLPAGVSGDFNGASENQSQRRNKILLENDGTLVLLRSGTNGSVEGLFYSYLANALSVTDMSRTVNTSRQYKPGYFGANRVAIGLYNTDSKILLGRYNDTSTNTNGLFISVTNGTLDDTQHSGLFVNQADISPEGGIEYAMMAEDGSVYVFNVLNENNKLQIVVVRVVLNIANGTYTATRITGWTTKTFYNTTYTAQNNIIITDHVNSTNIADKPYVYVPQNLAGMGIYMTSIDLYVAQEPGTQNFRFRMNGDAWFTMANYNTRPQHSYSFNFNMSTKQCNLDNGNDIGTFAAPFVVTNTGSALVATGNVLNKDPLYDHSGARNIYSSYYYLNTGEVFCTACPNLAEPLLMQKAKFTAQSIYSMINVRANISTDFRNGIVRPSFGSAVGSFINGVELLPGNSTKQYSRSTGSGAFRPSYAVHKSTPNFTFNSLSRGTIQGYEPTTERANVTDNTNNRVFISSISGNNVTTNGGIFMVGQRYSQSLSYDKNMNGTGIISVTQSVLDNFRNAEYAKVLSDWNLSSAAYKDMTLFVPQQTDIPAFVMLSTVTNTYQNYIRIAEVNVNTRSGNITTITFKRLVLENVYGSNFTPNGNGSFASSSVGLTIYDGGSFYFIGGSDPLSHPTIGNTNTHQFRAYVTKATAQFDSFIISGSHASYEQGIQPAALPGIGFGYFQAIDYVNKMTFQKCGTTIADYNNWTAQGAPINVVSQDVAQGFIVYFTEETPVILSGKAFTLPITSINLTTVKANPANTLFHVYVKMEQGNAKYVITTDVISETGSTAYNIFWIGTIQTNSLQISNVNIQKRSRLDIFGASLEAAGSSFPVSYGLPSQTGTINW
Physico‐chemical
properties
protein length:898 AA
molecular weight: 98240,42460 Da
isoelectric point:6,47762
aromaticity:0,11247
hydropathy:-0,21002

Domains

Domains [InterPro]

No domain annotations available.

Legend: Pfam SMART CDD TIGRFAM HAMAP SUPFAM PRINTS Gene3D PANTHER Other

Taxonomy

  Name Taxonomy ID Lineage
Phage Escherichia phage Ecwhy_1
[NCBI]
2419744 Viruses > Duplodnaviria > Heunggongvirae > Uroviricota > Caudoviricetes
Host No host information

Coding sequence (CDS)

Coding sequence (CDS)
Genbank protein accession
QAY00559.1 [NCBI]
Genbank nucleotide accession
MH791411 [NCBI]
CDS location
range 208568 -> 211264
strand -
CDS
ATGTCAAAGAAAATATATGGTGGCTTAGACGTCACTCATAATGTGACAGTTAAAGGGAAAAGAACAGTACAGGGCATCAATGATCAAGTATTTGATGCTAATGGTAATATGGAAATTGATACATATTCCAAAGAAGAATTACTTGATGTAATTAGTTTACTTCCAATCTCGCATTATGGTACACATAACTATCTACCTGCTGGGGTATCTGGTGATTTTAATGGTGCATCAGAAAACCAATCCCAACGTAGAAACAAAATACTTTTAGAAAACGATGGCACATTAGTACTATTACGTAGTGGTACAAACGGTAGTGTTGAAGGTTTATTTTATTCTTACCTAGCGAACGCATTATCCGTTACCGATATGAGCCGTACTGTTAATACAAGCCGTCAGTATAAGCCTGGCTATTTTGGGGCAAACCGTGTTGCTATTGGTTTATACAATACTGATTCAAAAATACTATTAGGTCGTTACAATGACACATCTACCAATACTAATGGGTTGTTTATATCAGTAACGAATGGTACATTGGATGATACACAACATAGTGGTCTTTTTGTGAACCAAGCTGATATTTCACCTGAAGGTGGGATTGAGTATGCTATGATGGCAGAAGACGGTTCAGTTTATGTATTCAACGTCCTTAATGAAAATAACAAACTACAGATTGTTGTTGTTCGTGTTGTATTGAATATTGCAAACGGTACTTATACTGCTACCCGAATCACGGGATGGACAACAAAAACTTTTTACAATACTACATATACTGCTCAGAATAATATAATCATTACTGACCATGTTAACTCAACTAATATTGCTGATAAGCCATATGTCTATGTACCACAAAATTTGGCTGGTATGGGGATTTATATGACATCTATCGATTTATATGTTGCTCAGGAGCCAGGAACGCAGAACTTCCGATTCAGAATGAATGGTGACGCATGGTTCACTATGGCAAACTACAATACCCGACCACAGCATAGTTATAGTTTCAATTTTAACATGAGTACTAAACAATGTAACTTGGATAATGGGAACGATATTGGTACATTTGCTGCACCATTTGTTGTTACAAACACAGGTTCGGCATTGGTTGCTACTGGTAACGTTCTCAACAAAGATCCTTTATATGACCATAGCGGTGCAAGAAACATTTACAGTTCATATTATTATCTGAATACAGGTGAAGTATTCTGTACTGCATGTCCAAACTTAGCAGAACCATTGCTTATGCAGAAAGCTAAATTTACTGCACAATCAATATATAGCATGATTAATGTACGTGCGAACATTTCAACTGATTTCAGAAATGGCATAGTTCGTCCTTCATTTGGTTCAGCAGTTGGTTCTTTTATTAACGGCGTAGAATTGCTTCCAGGGAACAGTACTAAACAGTATTCTCGTTCAACTGGTTCTGGTGCATTCCGTCCATCTTATGCAGTTCATAAAAGTACTCCTAACTTTACGTTTAATTCTTTATCACGTGGAACCATACAAGGTTATGAGCCAACAACAGAACGTGCTAATGTGACAGATAACACGAACAACAGAGTTTTTATTAGTTCCATTAGTGGTAATAATGTTACTACAAATGGTGGGATATTCATGGTAGGTCAGAGATATTCCCAATCACTATCATATGATAAAAATATGAACGGAACTGGTATCATTAGTGTTACTCAATCTGTACTTGATAATTTCAGAAATGCTGAATACGCCAAAGTACTAAGCGACTGGAACTTATCTTCCGCTGCTTATAAAGATATGACATTGTTTGTTCCACAACAAACTGATATACCTGCATTCGTAATGCTAAGTACAGTTACAAATACATATCAGAACTATATTAGAATTGCTGAAGTGAATGTGAATACACGTTCAGGTAATATCACTACAATTACCTTCAAGCGTTTAGTACTAGAAAACGTTTACGGAAGTAACTTTACACCAAATGGTAATGGTAGTTTTGCTTCGTCTTCTGTTGGTTTAACAATATATGATGGTGGTTCTTTCTACTTCATAGGTGGTTCTGACCCATTATCTCACCCAACTATTGGTAATACAAACACACACCAATTCAGAGCATATGTAACAAAAGCAACTGCACAATTTGATAGTTTTATTATCAGTGGTTCACATGCAAGTTATGAGCAAGGTATTCAACCAGCAGCTTTACCTGGTATTGGTTTTGGTTATTTCCAAGCAATAGACTATGTAAACAAAATGACTTTCCAAAAATGTGGTACTACAATTGCTGATTATAATAACTGGACTGCTCAAGGTGCACCGATTAACGTAGTTTCACAAGACGTTGCTCAAGGTTTTATTGTGTACTTCACAGAAGAAACTCCAGTAATTCTTTCTGGTAAAGCATTTACATTACCGATTACTAGTATAAACTTAACTACAGTTAAGGCTAATCCAGCAAACACATTATTTCATGTATATGTTAAAATGGAACAAGGTAATGCAAAATATGTAATAACAACTGATGTTATTTCAGAAACTGGTTCTACAGCATATAATATTTTCTGGATAGGTACTATACAAACAAACTCACTACAAATATCAAATGTTAACATACAAAAACGTTCAAGATTGGATATTTTCGGTGCGTCTCTTGAAGCTGCTGGTTCTTCATTCCCAGTTAGTTATGGTCTACCGTCACAAACAGGAACAATCAATTGGTAA

Gene Ontology

No Gene Ontology terms available.

Enzymatic activity

No enzymatic activity data available.

Tertiary structure

PDB ID
7d21c79c38a14faa13322051211350903706786ee40a4f904120dd07e779f1a7
ESMFold
Source ESMFold
Method ESMFold
Resolution 0,3800
Evidence 0,3800

Literature

Title Authors Date PMID Source
Ecwhy_1, Complete genome sequences of 3 novel enterobacteria, Pakpunavirus like phages Yuan,S., Ma,Y. and Liu,Q. 2020-07-27 GenBank