Detailed information
Overview
| Name | comEC | Type | Machinery gene |
| Locus tag | C4E16_RS06665 | Genome accession | NZ_CP026647 |
| Coordinates | 1478813..1480942 (-) | Length | 709 a.a. |
| NCBI ID | WP_001911453.1 | Uniprot ID | - |
| Organism | Vibrio cholerae O1 biovar El Tor strain HC1037 | ||
| Function | ssDNA transport through the inner membrane (predicted from homology) DNA binding and uptake |
||
Genomic Context
Location: 1473813..1485942
| Locus tag | Gene name | Coordinates (strand) | Size (bp) | Protein ID | Product | Description |
|---|---|---|---|---|---|---|
| C4E16_RS06645 (C4E16_06645) | kdsB | 1475106..1475864 (-) | 759 | WP_000011329.1 | 3-deoxy-manno-octulosonate cytidylyltransferase | - |
| C4E16_RS06650 (C4E16_06650) | - | 1475864..1476043 (-) | 180 | WP_000350068.1 | Trm112 family protein | - |
| C4E16_RS06655 (C4E16_06655) | lpxK | 1476024..1477031 (-) | 1008 | WP_001918694.1 | tetraacyldisaccharide 4'-kinase | - |
| C4E16_RS06660 (C4E16_06660) | msbA | 1477034..1478782 (-) | 1749 | WP_000052153.1 | lipid A ABC transporter ATP-binding protein/permease MsbA | - |
| C4E16_RS06665 (C4E16_06665) | comEC | 1478813..1480942 (-) | 2130 | WP_001911453.1 | DNA internalization-related competence protein ComEC/Rec2 | Machinery gene |
| C4E16_RS06670 (C4E16_06670) | - | 1481065..1481595 (+) | 531 | WP_001881633.1 | DUF2062 domain-containing protein | - |
| C4E16_RS06675 (C4E16_06675) | lolE | 1481710..1482954 (-) | 1245 | WP_000493010.1 | lipoprotein-releasing ABC transporter permease subunit LolE | - |
| C4E16_RS06680 (C4E16_06680) | lolD | 1482955..1483641 (-) | 687 | WP_001061290.1 | lipoprotein-releasing ABC transporter ATP-binding protein LolD | - |
| C4E16_RS06685 (C4E16_06685) | lolC | 1483634..1484842 (-) | 1209 | WP_000468900.1 | lipoprotein-releasing ABC transporter permease subunit LolC | - |
| C4E16_RS06690 (C4E16_06690) | - | 1485018..1485593 (+) | 576 | WP_000999601.1 | PilZ domain-containing protein | - |
Sequence
Protein
Download Length: 709 a.a. Molecular weight: 80009.33 Da Isoelectric Point: 7.7984
>NTDB_id=271179 C4E16_RS06665 WP_001911453.1 1478813..1480942(-) (comEC) [Vibrio cholerae O1 biovar El Tor strain HC1037]
MVLLGYHRVGRQFLGFVAAILTIVLQGNLIRDQSNVLYQAGPDIIIKGRVDSFFTQTRYAYEGFVLIHEVNGQTLNKMTR
PRIRLSAPLLLQPNDRVEFSVTLKPIVGRLNQTGFDLEAHYMAQSVVARAVVKPDTAYQIVQESGIRSSLFFELEQLTHT
SPYQGLILALTFGERKGIDEQEWQALRNSGLIHLVAISGLHIGIAFSVGYFLGLGMMRFHAQLLWSPFVCGALLAVLYAW
LAGFTLPTQRALIMCLLNVALIMLAFPLSALKRILLTLVAVLLWSPFASLSNSFWMSFLAVAIVLYQLASQSQRQVWWKA
LLWAQVFLVCLMAPVTAYFFGGLSVTAVLYNLVFIPWFSLVIVPALFLGLLLMVVWPSVAAAYWPWVDWTFLPLDWALQF
ADVGWWVVPSKVQGVVAASVAILLLYRFMSLKACSLLLGMIGLWWWFPSLTPLWRMDVLDVGHGLAIVIEQDERAIVYDT
GSSWPGGSYVQSVIEPMLQQRGLRQVDGVILSHLDNDHAGDWQGLAERWQPNWIRASQLGTEFMPCIRGESWQWQSLHFT
VLWPPQAVSRAYNQHSCVIRMTDTQSNHSVLLSGDVTAMGEWLLARDGAQLQSEVMIVPHHGSKTSSTAEFIAQVNPKLA
IASVAKDNRWNLPNPQVVARYQAQQVEWLDTGHAGQISLFFYLDQLDWFTQRSLGWQPWYRQMLRKGVE
MVLLGYHRVGRQFLGFVAAILTIVLQGNLIRDQSNVLYQAGPDIIIKGRVDSFFTQTRYAYEGFVLIHEVNGQTLNKMTR
PRIRLSAPLLLQPNDRVEFSVTLKPIVGRLNQTGFDLEAHYMAQSVVARAVVKPDTAYQIVQESGIRSSLFFELEQLTHT
SPYQGLILALTFGERKGIDEQEWQALRNSGLIHLVAISGLHIGIAFSVGYFLGLGMMRFHAQLLWSPFVCGALLAVLYAW
LAGFTLPTQRALIMCLLNVALIMLAFPLSALKRILLTLVAVLLWSPFASLSNSFWMSFLAVAIVLYQLASQSQRQVWWKA
LLWAQVFLVCLMAPVTAYFFGGLSVTAVLYNLVFIPWFSLVIVPALFLGLLLMVVWPSVAAAYWPWVDWTFLPLDWALQF
ADVGWWVVPSKVQGVVAASVAILLLYRFMSLKACSLLLGMIGLWWWFPSLTPLWRMDVLDVGHGLAIVIEQDERAIVYDT
GSSWPGGSYVQSVIEPMLQQRGLRQVDGVILSHLDNDHAGDWQGLAERWQPNWIRASQLGTEFMPCIRGESWQWQSLHFT
VLWPPQAVSRAYNQHSCVIRMTDTQSNHSVLLSGDVTAMGEWLLARDGAQLQSEVMIVPHHGSKTSSTAEFIAQVNPKLA
IASVAKDNRWNLPNPQVVARYQAQQVEWLDTGHAGQISLFFYLDQLDWFTQRSLGWQPWYRQMLRKGVE
Nucleotide
Download Length: 2130 bp
>NTDB_id=271179 C4E16_RS06665 WP_001911453.1 1478813..1480942(-) (comEC) [Vibrio cholerae O1 biovar El Tor strain HC1037]
ATGGTTTTGCTCGGTTATCACCGAGTTGGCCGTCAATTCCTTGGCTTCGTGGCTGCCATACTAACCATTGTGCTACAGGG
CAACCTTATACGAGATCAATCCAATGTGCTCTATCAAGCAGGGCCGGATATTATCATAAAAGGCCGTGTTGACAGCTTTT
TTACGCAAACTCGTTACGCTTATGAGGGTTTTGTCCTCATTCATGAAGTGAATGGACAAACCTTAAACAAAATGACTCGC
CCTCGCATACGTTTAAGTGCCCCTTTACTGTTACAACCCAATGATCGCGTCGAATTTTCGGTAACTCTCAAGCCGATAGT
GGGTCGACTCAACCAAACCGGCTTTGATTTAGAAGCGCATTACATGGCGCAATCTGTCGTCGCACGAGCGGTCGTAAAAC
CTGACACTGCTTATCAAATTGTGCAAGAGAGTGGCATAAGGTCAAGTTTGTTTTTTGAGCTAGAGCAATTAACGCATACC
AGCCCATACCAAGGATTGATCTTAGCCCTGACGTTTGGCGAGCGAAAAGGTATTGATGAGCAAGAGTGGCAAGCCTTACG
CAATAGTGGCTTAATTCATTTAGTGGCCATTTCGGGGCTGCACATTGGTATCGCTTTTAGCGTGGGGTATTTTCTCGGGC
TCGGCATGATGCGTTTTCATGCTCAGTTATTGTGGTCCCCTTTTGTGTGTGGGGCTTTACTGGCGGTGCTCTACGCTTGG
CTGGCCGGATTTACGTTGCCTACTCAGCGTGCATTGATTATGTGCTTACTCAATGTGGCGTTGATCATGTTGGCTTTTCC
TCTTTCCGCGCTCAAGCGGATTCTACTCACCTTAGTCGCGGTCTTGCTTTGGTCGCCATTCGCCTCACTTTCAAACAGTT
TCTGGATGTCGTTTTTGGCGGTCGCGATTGTTCTCTACCAATTAGCCAGTCAAAGCCAGCGTCAGGTGTGGTGGAAAGCT
CTTCTTTGGGCGCAGGTGTTCCTCGTCTGTTTAATGGCACCGGTCACGGCCTATTTTTTCGGTGGCTTAAGCGTAACGGC
AGTTCTGTACAATTTGGTGTTTATTCCTTGGTTTTCGTTGGTGATTGTCCCAGCTTTGTTTTTGGGTCTATTACTCATGG
TGGTATGGCCTAGTGTGGCCGCCGCTTACTGGCCTTGGGTGGATTGGACGTTTTTACCGCTCGATTGGGCTTTGCAGTTT
GCCGATGTAGGCTGGTGGGTGGTCCCCAGCAAAGTACAAGGTGTGGTCGCAGCGAGTGTGGCCATCCTCTTGCTTTATCG
ATTTATGAGCCTAAAAGCCTGCAGCTTATTATTGGGTATGATTGGCTTATGGTGGTGGTTTCCCTCTCTCACTCCACTTT
GGCGAATGGATGTGCTGGATGTTGGACATGGCTTGGCGATTGTGATTGAGCAAGATGAGCGAGCAATTGTCTACGATACA
GGCAGCAGTTGGCCGGGAGGCAGCTATGTGCAAAGCGTGATTGAGCCTATGCTCCAACAGCGGGGGCTACGCCAAGTGGA
TGGAGTGATTTTAAGTCATCTTGATAATGATCATGCGGGTGATTGGCAAGGTTTAGCTGAGCGCTGGCAACCCAATTGGA
TTCGTGCCAGCCAACTCGGGACAGAGTTTATGCCTTGTATCCGTGGTGAAAGCTGGCAGTGGCAATCTCTCCATTTTACG
GTGTTATGGCCACCACAAGCGGTTAGCCGAGCGTACAACCAGCATTCGTGTGTGATTCGTATGACCGATACTCAGTCTAA
CCATTCTGTACTGCTCTCCGGGGATGTCACAGCCATGGGGGAGTGGCTGCTTGCTCGCGACGGAGCGCAACTGCAAAGTG
AGGTGATGATCGTGCCGCACCACGGCAGTAAAACATCGTCCACCGCAGAGTTTATTGCCCAAGTGAATCCCAAACTTGCG
ATTGCTTCTGTGGCGAAAGATAACCGCTGGAATTTGCCTAATCCGCAAGTCGTGGCACGTTATCAAGCTCAGCAAGTTGA
GTGGCTAGATACTGGACACGCTGGGCAAATTAGCCTCTTTTTCTATCTAGATCAGCTGGATTGGTTTACCCAGCGTAGCC
TTGGCTGGCAGCCTTGGTATAGGCAGATGCTGCGTAAAGGAGTAGAATGA
ATGGTTTTGCTCGGTTATCACCGAGTTGGCCGTCAATTCCTTGGCTTCGTGGCTGCCATACTAACCATTGTGCTACAGGG
CAACCTTATACGAGATCAATCCAATGTGCTCTATCAAGCAGGGCCGGATATTATCATAAAAGGCCGTGTTGACAGCTTTT
TTACGCAAACTCGTTACGCTTATGAGGGTTTTGTCCTCATTCATGAAGTGAATGGACAAACCTTAAACAAAATGACTCGC
CCTCGCATACGTTTAAGTGCCCCTTTACTGTTACAACCCAATGATCGCGTCGAATTTTCGGTAACTCTCAAGCCGATAGT
GGGTCGACTCAACCAAACCGGCTTTGATTTAGAAGCGCATTACATGGCGCAATCTGTCGTCGCACGAGCGGTCGTAAAAC
CTGACACTGCTTATCAAATTGTGCAAGAGAGTGGCATAAGGTCAAGTTTGTTTTTTGAGCTAGAGCAATTAACGCATACC
AGCCCATACCAAGGATTGATCTTAGCCCTGACGTTTGGCGAGCGAAAAGGTATTGATGAGCAAGAGTGGCAAGCCTTACG
CAATAGTGGCTTAATTCATTTAGTGGCCATTTCGGGGCTGCACATTGGTATCGCTTTTAGCGTGGGGTATTTTCTCGGGC
TCGGCATGATGCGTTTTCATGCTCAGTTATTGTGGTCCCCTTTTGTGTGTGGGGCTTTACTGGCGGTGCTCTACGCTTGG
CTGGCCGGATTTACGTTGCCTACTCAGCGTGCATTGATTATGTGCTTACTCAATGTGGCGTTGATCATGTTGGCTTTTCC
TCTTTCCGCGCTCAAGCGGATTCTACTCACCTTAGTCGCGGTCTTGCTTTGGTCGCCATTCGCCTCACTTTCAAACAGTT
TCTGGATGTCGTTTTTGGCGGTCGCGATTGTTCTCTACCAATTAGCCAGTCAAAGCCAGCGTCAGGTGTGGTGGAAAGCT
CTTCTTTGGGCGCAGGTGTTCCTCGTCTGTTTAATGGCACCGGTCACGGCCTATTTTTTCGGTGGCTTAAGCGTAACGGC
AGTTCTGTACAATTTGGTGTTTATTCCTTGGTTTTCGTTGGTGATTGTCCCAGCTTTGTTTTTGGGTCTATTACTCATGG
TGGTATGGCCTAGTGTGGCCGCCGCTTACTGGCCTTGGGTGGATTGGACGTTTTTACCGCTCGATTGGGCTTTGCAGTTT
GCCGATGTAGGCTGGTGGGTGGTCCCCAGCAAAGTACAAGGTGTGGTCGCAGCGAGTGTGGCCATCCTCTTGCTTTATCG
ATTTATGAGCCTAAAAGCCTGCAGCTTATTATTGGGTATGATTGGCTTATGGTGGTGGTTTCCCTCTCTCACTCCACTTT
GGCGAATGGATGTGCTGGATGTTGGACATGGCTTGGCGATTGTGATTGAGCAAGATGAGCGAGCAATTGTCTACGATACA
GGCAGCAGTTGGCCGGGAGGCAGCTATGTGCAAAGCGTGATTGAGCCTATGCTCCAACAGCGGGGGCTACGCCAAGTGGA
TGGAGTGATTTTAAGTCATCTTGATAATGATCATGCGGGTGATTGGCAAGGTTTAGCTGAGCGCTGGCAACCCAATTGGA
TTCGTGCCAGCCAACTCGGGACAGAGTTTATGCCTTGTATCCGTGGTGAAAGCTGGCAGTGGCAATCTCTCCATTTTACG
GTGTTATGGCCACCACAAGCGGTTAGCCGAGCGTACAACCAGCATTCGTGTGTGATTCGTATGACCGATACTCAGTCTAA
CCATTCTGTACTGCTCTCCGGGGATGTCACAGCCATGGGGGAGTGGCTGCTTGCTCGCGACGGAGCGCAACTGCAAAGTG
AGGTGATGATCGTGCCGCACCACGGCAGTAAAACATCGTCCACCGCAGAGTTTATTGCCCAAGTGAATCCCAAACTTGCG
ATTGCTTCTGTGGCGAAAGATAACCGCTGGAATTTGCCTAATCCGCAAGTCGTGGCACGTTATCAAGCTCAGCAAGTTGA
GTGGCTAGATACTGGACACGCTGGGCAAATTAGCCTCTTTTTCTATCTAGATCAGCTGGATTGGTTTACCCAGCGTAGCC
TTGGCTGGCAGCCTTGGTATAGGCAGATGCTGCGTAAAGGAGTAGAATGA
3D structure
| Source | ID | Structure |
|---|
Similar proteins
Only experimentally validated proteins are listed.
| Protein | Organism | Identities (%) | Coverage (%) | Ha-value |
|---|---|---|---|---|
| comEC | Vibrio cholerae strain A1552 |
100 |
100 |
1 |
| comEC | Vibrio parahaemolyticus RIMD 2210633 |
41.433 |
100 |
0.416 |
| comEC | Vibrio campbellii strain DS40M4 |
41.301 |
99.718 |
0.412 |