Protein

Genbank accession
AYN56439.1 [GenBank]
Protein name
long-tail fiber proximal subunit
RBP type
TF
Evidence RBPdetect
Probability 0,90
TF
Evidence RBPdetect2
Probability 0,77
Protein sequence
MRKGQTVKIKAAEGDTIASSVALLQFPKRSEYPPDAQWVSVTELEFNGTTSYVPVLELAYIEDTVAGTRYWVVQQNVPTVERVDAGTDTTRARVGVIALATQAQANVDFENSPAKEVAITPETLANRTATEARRGIAKIATTAQVNQNSTATFVDDTIVTPKKLNERTATETRRGLAEIATQVETDAGLDDTTIITPKKLQARQGSETLSGIVKYVSTTSATPAETRGAAGTNVYNKTVNNLTISPKALDQYKATYAQQGAVILAVDSEVIAGQSQAGYSHAVVTPETLHKKTSTDGRIGLIEIATQAETNAGTDYTRAVTPKTLNDRKATEGLSGIAELATQVEFDTGTDDTRISTPLKIKTHFDSSDRTSVNSDSGLIEEGTLWNHYTLDISKANETQRGTLRVATQAESNAGTLDDVLITPKKLLGTKSTETSEGVIKVATQAETVTGTSANTAVSPKNLKWIVQNEPTWAATTAIRGFVKTSSGSITFVGNDTVGSTQPLESYEKNSYAVSPYELNRVLANYLPLKAKAVDSNLLDGLDSSQFIRRDIAQTVNGSLTLTQQTNLSAPLVSTSTATFGGSVSANSTLTISNTGTTSSRFTFEKGPASGSNADSALYVRVWGNKYSGGSDVTRATIIEFSDATGSHFYSQRDTSNNVLFNISGTMQSVNASVRGVLNVTGVSTFNSSVTANGEFISKSANAFRAISGDYGFFIRNDAVNTYFMLTASGDQTGGFNGLRPLAINNASGQVTIGESLIIAKGATINSGGLTVNSRIRSQGSKTADLYTRAPTSDTVGFWSIDINDSATYNQFPGYFKMVEKTNEVTGLPYLERGEEVKSPGTLTQFGNTLDSLYQDWITYPTTPEARTTRWTRTWQKTKNSWSSFVQVFDGGNPPQPSDIGALPSDNATIGNLTIRDFLRIGNVRIIPDPVNKSVKFEWIE
Physico‐chemical
properties
protein length:941 AA
molecular weight: 100941,65830 Da
isoelectric point:5,52602
aromaticity:0,07226
hydropathy:-0,34240

Domains

Domains [InterPro]
AYN56439.1
1 941
Legend: Pfam SMART CDD TIGRFAM HAMAP SUPFAM PRINTS Gene3D PANTHER Other

Taxonomy

  Name Taxonomy ID Lineage
Phage Escherichia phage p000v
[NCBI]
2479933 Viruses > Duplodnaviria > Heunggongvirae > Uroviricota > Caudoviricetes
Host No host information

Coding sequence (CDS)

Coding sequence (CDS)
Genbank protein accession
AYN56439.1 [NCBI]
Genbank nucleotide accession
MK047717 [NCBI]
CDS location
range 164810 -> 167635
strand -
CDS
ATGCGTAAAGGCCAAACAGTAAAAATTAAAGCTGCTGAAGGTGATACAATTGCTTCTTCTGTTGCATTGCTTCAGTTCCCTAAACGTTCTGAATATCCACCTGATGCTCAATGGGTCTCTGTTACTGAATTGGAATTCAATGGCACTACTTCATATGTTCCTGTATTAGAATTAGCGTATATCGAAGACACCGTTGCTGGTACACGTTATTGGGTTGTTCAACAGAATGTCCCAACTGTAGAACGTGTCGATGCTGGAACCGATACAACTAGAGCTCGTGTAGGTGTTATTGCTCTTGCAACTCAGGCACAGGCAAATGTTGATTTTGAAAATTCTCCGGCTAAAGAAGTTGCTATTACTCCTGAAACATTGGCAAATCGTACAGCAACTGAAGCTCGACGTGGTATCGCTAAAATTGCTACAACAGCACAAGTTAACCAGAATTCCACTGCAACATTTGTGGATGATACTATTGTTACGCCTAAAAAACTAAATGAGCGTACAGCAACTGAAACTCGTCGTGGTCTTGCTGAAATTGCTACTCAGGTTGAAACCGATGCTGGTCTTGATGACACAACGATTATTACACCTAAGAAATTACAGGCACGTCAGGGTTCTGAAACACTGTCTGGTATAGTTAAATATGTATCAACTACTTCTGCTACTCCTGCTGAAACTCGTGGGGCTGCAGGCACTAACGTTTATAATAAAACCGTAAATAATTTAACTATTTCTCCTAAAGCCCTTGACCAATATAAAGCAACTTATGCTCAACAAGGTGCAGTAATTTTAGCTGTTGATAGTGAAGTAATTGCTGGTCAATCACAAGCAGGTTATTCTCACGCTGTAGTAACTCCTGAAACACTACATAAGAAAACTTCTACTGATGGACGTATTGGTTTAATTGAAATTGCTACGCAAGCAGAAACTAATGCTGGGACTGATTATACACGTGCAGTAACGCCTAAGACGTTAAATGATAGGAAAGCTACGGAAGGATTATCCGGCATAGCCGAACTTGCTACGCAAGTTGAATTTGATACTGGAACTGATGATACTCGTATCTCGACTCCACTGAAAATTAAAACTCATTTTGATTCTTCTGACCGTACCAGTGTTAATTCTGATTCCGGACTTATTGAAGAAGGAACCTTGTGGAACCATTATACTCTTGATATTTCTAAAGCAAATGAAACACAACGTGGTACACTCCGCGTAGCGACCCAGGCAGAATCTAATGCAGGAACTTTAGATGATGTTCTTATTACTCCTAAAAAGCTTTTAGGGACTAAGTCCACTGAAACGTCTGAAGGCGTAATTAAGGTTGCTACTCAGGCTGAAACTGTAACAGGAACTTCTGCTAATACTGCTGTATCTCCTAAGAATTTAAAATGGATTGTCCAGAATGAACCAACATGGGCTGCTACTACAGCAATTCGCGGATTCGTTAAAACTTCGTCCGGTTCTATTACATTTGTTGGTAATGATACAGTTGGTTCAACACAACCTTTAGAATCATACGAAAAAAATAGCTATGCAGTATCACCATATGAATTAAACCGTGTACTTGCTAACTATTTGCCATTGAAAGCTAAAGCCGTAGATAGTAATTTATTAGATGGCCTAGATTCATCTCAGTTCATTCGTAGGGACATTGCACAGACGGTTAATGGTTCACTAACCTTAACCCAACAAACGAATCTGAGTGCCCCTCTTGTATCAACTAGTACTGCTACGTTTGGTGGTTCAGTATCTGCTAATAGTACACTGACTATTTCTAATACTGGAACGACTTCTTCTCGATTTACATTTGAGAAAGGTCCTGCTTCTGGTAGTAATGCTGATTCTGCATTGTATGTTCGTGTATGGGGTAATAAGTACAGCGGAGGTTCTGATGTAACTCGTGCAACGATTATAGAATTCTCTGATGCTACCGGCTCTCATTTCTATTCTCAAAGAGATACGTCGAATAATGTGCTGTTCAATATTTCAGGTACGATGCAATCAGTCAACGCTAGCGTTCGTGGCGTTCTGAACGTTACAGGTGTCTCAACGTTTAATAGTTCAGTTACAGCCAATGGTGAATTCATTAGTAAGTCTGCAAATGCTTTCAGAGCAATTAGTGGTGATTACGGATTCTTTATTCGTAATGACGCTGTTAATACCTATTTTATGCTCACTGCATCAGGCGATCAGACCGGCGGATTTAATGGATTACGTCCTTTAGCTATTAATAATGCATCTGGCCAAGTAACGATTGGTGAAAGCTTAATCATTGCCAAAGGTGCTACTATAAATTCAGGTGGTTTAACTGTTAACTCGAGAATTCGTTCTCAGGGCTCTAAAACTGCTGATTTATACACTCGCGCACCGACATCTGATACTGTAGGATTCTGGTCAATCGATATTAACGATTCAGCCACTTATAACCAGTTCCCGGGTTATTTTAAAATGGTTGAAAAAACTAATGAAGTGACTGGGCTTCCATACTTAGAACGTGGCGAAGAAGTTAAATCTCCTGGTACATTGACTCAGTTTGGTAATACGCTTGATTCTCTTTACCAAGATTGGATTACTTATCCAACAACTCCAGAAGCACGTACAACCCGTTGGACTCGTACATGGCAGAAAACTAAAAATTCTTGGTCAAGTTTTGTTCAGGTATTTGATGGCGGAAACCCTCCTCAACCTTCTGATATTGGTGCTTTACCTTCTGATAATGCAACAATCGGAAACTTGACAATAAGGGATTTCTTAAGGATTGGTAATGTCCGCATTATTCCAGACCCTGTGAATAAATCTGTTAAATTCGAGTGGATTGAATAA

Gene Ontology

No Gene Ontology terms available.

Enzymatic activity

No enzymatic activity data available.

Tertiary structure

PDB ID
59e341698c4df9ed2fb8d8e2093f2b6b6301cadda577edff5b6e149f811f5694
ESMFold
Source ESMFold
Method ESMFold
Resolution 0,5697
Evidence 0,5697

Literature

Title Authors Date PMID Source
Whole-genome sequence of phages p000v and p000y infecting the bacterial pathogen Shigatoxigenic Escherichia coli Howard-Varona,C., Vik,D.R., Solonenko,N.E., Li,Y.-F., Gazitua,M.C., Hobbs,Z., Honaker,R.W., Kinkhabwala,A.A. and Sullivan,M.B. 2018-11-21 GenBank