Detailed information
Overview
| Name | comEC | Type | Machinery gene |
| Locus tag | GQX72_RS04965 | Genome accession | NZ_CP047059 |
| Coordinates | 1052455..1054584 (+) | Length | 709 a.a. |
| NCBI ID | WP_001911453.1 | Uniprot ID | - |
| Organism | Vibrio cholerae isolate CTMA_1441 | ||
| Function | ssDNA transport through the inner membrane (predicted from homology) DNA binding and uptake |
||
Genomic Context
Location: 1047455..1059584
| Locus tag | Gene name | Coordinates (strand) | Size (bp) | Protein ID | Product | Description |
|---|---|---|---|---|---|---|
| GQX72_RS04935 (GQX72_04935) | - | 1047804..1048379 (-) | 576 | WP_000999601.1 | PilZ domain-containing protein | - |
| GQX72_RS04940 (GQX72_04940) | lolC | 1048555..1049763 (+) | 1209 | WP_000468900.1 | lipoprotein-releasing ABC transporter permease subunit LolC | - |
| GQX72_RS04945 (GQX72_04945) | lolD | 1049756..1050442 (+) | 687 | WP_001061290.1 | lipoprotein-releasing ABC transporter ATP-binding protein LolD | - |
| GQX72_RS04950 (GQX72_04950) | lolE | 1050443..1051687 (+) | 1245 | WP_000493010.1 | lipoprotein-releasing ABC transporter permease subunit LolE | - |
| GQX72_RS04960 (GQX72_04960) | - | 1051802..1052332 (-) | 531 | WP_001881633.1 | DUF2062 domain-containing protein | - |
| GQX72_RS04965 (GQX72_04965) | comEC | 1052455..1054584 (+) | 2130 | WP_001911453.1 | DNA internalization-related competence protein ComEC/Rec2 | Machinery gene |
| GQX72_RS04970 (GQX72_04970) | msbA | 1054615..1056363 (+) | 1749 | WP_000052153.1 | lipid A ABC transporter ATP-binding protein/permease MsbA | - |
| GQX72_RS04975 (GQX72_04975) | lpxK | 1056366..1057373 (+) | 1008 | WP_001918694.1 | tetraacyldisaccharide 4'-kinase | - |
| GQX72_RS04980 (GQX72_04980) | - | 1057354..1057533 (+) | 180 | WP_000350068.1 | Trm112 family protein | - |
| GQX72_RS04985 (GQX72_04985) | kdsB | 1057533..1058291 (+) | 759 | WP_000011329.1 | 3-deoxy-manno-octulosonate cytidylyltransferase | - |
Sequence
Protein
Download Length: 709 a.a. Molecular weight: 80009.33 Da Isoelectric Point: 7.7984
>NTDB_id=410117 GQX72_RS04965 WP_001911453.1 1052455..1054584(+) (comEC) [Vibrio cholerae isolate CTMA_1441]
MVLLGYHRVGRQFLGFVAAILTIVLQGNLIRDQSNVLYQAGPDIIIKGRVDSFFTQTRYAYEGFVLIHEVNGQTLNKMTR
PRIRLSAPLLLQPNDRVEFSVTLKPIVGRLNQTGFDLEAHYMAQSVVARAVVKPDTAYQIVQESGIRSSLFFELEQLTHT
SPYQGLILALTFGERKGIDEQEWQALRNSGLIHLVAISGLHIGIAFSVGYFLGLGMMRFHAQLLWSPFVCGALLAVLYAW
LAGFTLPTQRALIMCLLNVALIMLAFPLSALKRILLTLVAVLLWSPFASLSNSFWMSFLAVAIVLYQLASQSQRQVWWKA
LLWAQVFLVCLMAPVTAYFFGGLSVTAVLYNLVFIPWFSLVIVPALFLGLLLMVVWPSVAAAYWPWVDWTFLPLDWALQF
ADVGWWVVPSKVQGVVAASVAILLLYRFMSLKACSLLLGMIGLWWWFPSLTPLWRMDVLDVGHGLAIVIEQDERAIVYDT
GSSWPGGSYVQSVIEPMLQQRGLRQVDGVILSHLDNDHAGDWQGLAERWQPNWIRASQLGTEFMPCIRGESWQWQSLHFT
VLWPPQAVSRAYNQHSCVIRMTDTQSNHSVLLSGDVTAMGEWLLARDGAQLQSEVMIVPHHGSKTSSTAEFIAQVNPKLA
IASVAKDNRWNLPNPQVVARYQAQQVEWLDTGHAGQISLFFYLDQLDWFTQRSLGWQPWYRQMLRKGVE
MVLLGYHRVGRQFLGFVAAILTIVLQGNLIRDQSNVLYQAGPDIIIKGRVDSFFTQTRYAYEGFVLIHEVNGQTLNKMTR
PRIRLSAPLLLQPNDRVEFSVTLKPIVGRLNQTGFDLEAHYMAQSVVARAVVKPDTAYQIVQESGIRSSLFFELEQLTHT
SPYQGLILALTFGERKGIDEQEWQALRNSGLIHLVAISGLHIGIAFSVGYFLGLGMMRFHAQLLWSPFVCGALLAVLYAW
LAGFTLPTQRALIMCLLNVALIMLAFPLSALKRILLTLVAVLLWSPFASLSNSFWMSFLAVAIVLYQLASQSQRQVWWKA
LLWAQVFLVCLMAPVTAYFFGGLSVTAVLYNLVFIPWFSLVIVPALFLGLLLMVVWPSVAAAYWPWVDWTFLPLDWALQF
ADVGWWVVPSKVQGVVAASVAILLLYRFMSLKACSLLLGMIGLWWWFPSLTPLWRMDVLDVGHGLAIVIEQDERAIVYDT
GSSWPGGSYVQSVIEPMLQQRGLRQVDGVILSHLDNDHAGDWQGLAERWQPNWIRASQLGTEFMPCIRGESWQWQSLHFT
VLWPPQAVSRAYNQHSCVIRMTDTQSNHSVLLSGDVTAMGEWLLARDGAQLQSEVMIVPHHGSKTSSTAEFIAQVNPKLA
IASVAKDNRWNLPNPQVVARYQAQQVEWLDTGHAGQISLFFYLDQLDWFTQRSLGWQPWYRQMLRKGVE
Nucleotide
Download Length: 2130 bp
>NTDB_id=410117 GQX72_RS04965 WP_001911453.1 1052455..1054584(+) (comEC) [Vibrio cholerae isolate CTMA_1441]
ATGGTTTTGCTCGGTTATCACCGAGTTGGCCGTCAATTCCTTGGCTTCGTGGCTGCCATACTAACCATTGTGCTACAGGG
CAACCTTATACGAGATCAATCCAATGTGCTCTATCAAGCAGGGCCGGATATTATCATAAAAGGCCGTGTTGACAGCTTTT
TTACGCAAACTCGTTACGCTTATGAGGGTTTTGTCCTCATTCATGAAGTGAATGGACAAACCTTAAACAAAATGACTCGC
CCTCGCATACGTTTAAGTGCCCCTTTACTGTTACAACCCAATGATCGCGTCGAATTTTCGGTAACTCTCAAGCCGATAGT
GGGTCGACTCAACCAAACCGGCTTTGATTTAGAAGCGCATTACATGGCGCAATCTGTCGTCGCACGAGCGGTCGTAAAAC
CTGACACTGCTTATCAAATTGTGCAAGAGAGTGGCATAAGGTCAAGTTTGTTTTTTGAGCTAGAGCAATTAACGCATACC
AGCCCATACCAAGGATTGATCTTAGCCCTGACGTTTGGCGAGCGAAAAGGTATTGATGAGCAAGAGTGGCAAGCCTTACG
CAATAGTGGCTTAATTCATTTAGTGGCCATTTCGGGGCTGCACATTGGTATCGCTTTTAGCGTGGGGTATTTTCTCGGGC
TCGGCATGATGCGTTTTCATGCTCAGTTATTGTGGTCCCCTTTTGTGTGTGGGGCTTTACTGGCGGTGCTCTACGCTTGG
CTGGCCGGATTTACGTTGCCTACTCAGCGTGCATTGATTATGTGCTTACTCAATGTGGCGTTGATCATGTTGGCTTTTCC
TCTTTCCGCGCTCAAGCGGATTCTACTCACCTTAGTCGCGGTCTTGCTTTGGTCGCCATTCGCCTCACTTTCAAACAGTT
TCTGGATGTCGTTTTTGGCGGTCGCGATTGTTCTCTACCAATTAGCCAGTCAAAGCCAGCGTCAGGTGTGGTGGAAAGCT
CTTCTTTGGGCGCAGGTGTTCCTCGTCTGTTTAATGGCACCGGTCACGGCCTATTTTTTCGGTGGCTTAAGCGTAACGGC
AGTTCTGTACAATTTGGTGTTTATTCCTTGGTTTTCGTTGGTGATTGTCCCAGCTTTGTTTTTGGGTCTATTACTCATGG
TGGTATGGCCTAGTGTGGCCGCCGCTTACTGGCCTTGGGTGGATTGGACGTTTTTACCGCTCGATTGGGCTTTGCAGTTT
GCCGATGTAGGCTGGTGGGTGGTCCCCAGCAAAGTACAAGGTGTGGTCGCAGCGAGTGTGGCCATCCTCTTGCTTTATCG
ATTTATGAGCCTAAAAGCCTGCAGCTTATTATTGGGTATGATTGGCTTATGGTGGTGGTTTCCCTCTCTCACTCCACTTT
GGCGAATGGATGTGCTGGATGTTGGACATGGCTTGGCGATTGTGATTGAGCAAGATGAGCGAGCAATTGTCTACGATACA
GGCAGCAGTTGGCCGGGAGGCAGCTATGTGCAAAGCGTGATTGAGCCTATGCTCCAACAGCGGGGGCTACGCCAAGTGGA
TGGAGTGATTTTAAGTCATCTTGATAATGATCATGCGGGTGATTGGCAAGGTTTAGCTGAGCGCTGGCAACCCAATTGGA
TTCGTGCCAGCCAACTCGGGACAGAGTTTATGCCTTGTATCCGTGGTGAAAGCTGGCAGTGGCAATCTCTCCATTTTACG
GTGTTATGGCCACCACAAGCGGTTAGCCGAGCGTACAACCAGCATTCGTGTGTGATTCGTATGACCGATACTCAGTCTAA
CCATTCTGTACTGCTCTCCGGGGATGTCACAGCCATGGGGGAGTGGCTGCTTGCTCGCGACGGAGCGCAACTGCAAAGTG
AGGTGATGATCGTGCCGCACCACGGCAGTAAAACATCGTCCACCGCAGAGTTTATTGCCCAAGTGAATCCCAAACTTGCG
ATTGCTTCTGTGGCGAAAGATAACCGCTGGAATTTGCCTAATCCGCAAGTCGTGGCACGTTATCAAGCTCAGCAAGTTGA
GTGGCTAGATACTGGACACGCTGGGCAAATTAGCCTCTTTTTCTATCTAGATCAGCTGGATTGGTTTACCCAGCGTAGCC
TTGGCTGGCAGCCTTGGTATAGGCAGATGCTGCGTAAAGGAGTAGAATGA
ATGGTTTTGCTCGGTTATCACCGAGTTGGCCGTCAATTCCTTGGCTTCGTGGCTGCCATACTAACCATTGTGCTACAGGG
CAACCTTATACGAGATCAATCCAATGTGCTCTATCAAGCAGGGCCGGATATTATCATAAAAGGCCGTGTTGACAGCTTTT
TTACGCAAACTCGTTACGCTTATGAGGGTTTTGTCCTCATTCATGAAGTGAATGGACAAACCTTAAACAAAATGACTCGC
CCTCGCATACGTTTAAGTGCCCCTTTACTGTTACAACCCAATGATCGCGTCGAATTTTCGGTAACTCTCAAGCCGATAGT
GGGTCGACTCAACCAAACCGGCTTTGATTTAGAAGCGCATTACATGGCGCAATCTGTCGTCGCACGAGCGGTCGTAAAAC
CTGACACTGCTTATCAAATTGTGCAAGAGAGTGGCATAAGGTCAAGTTTGTTTTTTGAGCTAGAGCAATTAACGCATACC
AGCCCATACCAAGGATTGATCTTAGCCCTGACGTTTGGCGAGCGAAAAGGTATTGATGAGCAAGAGTGGCAAGCCTTACG
CAATAGTGGCTTAATTCATTTAGTGGCCATTTCGGGGCTGCACATTGGTATCGCTTTTAGCGTGGGGTATTTTCTCGGGC
TCGGCATGATGCGTTTTCATGCTCAGTTATTGTGGTCCCCTTTTGTGTGTGGGGCTTTACTGGCGGTGCTCTACGCTTGG
CTGGCCGGATTTACGTTGCCTACTCAGCGTGCATTGATTATGTGCTTACTCAATGTGGCGTTGATCATGTTGGCTTTTCC
TCTTTCCGCGCTCAAGCGGATTCTACTCACCTTAGTCGCGGTCTTGCTTTGGTCGCCATTCGCCTCACTTTCAAACAGTT
TCTGGATGTCGTTTTTGGCGGTCGCGATTGTTCTCTACCAATTAGCCAGTCAAAGCCAGCGTCAGGTGTGGTGGAAAGCT
CTTCTTTGGGCGCAGGTGTTCCTCGTCTGTTTAATGGCACCGGTCACGGCCTATTTTTTCGGTGGCTTAAGCGTAACGGC
AGTTCTGTACAATTTGGTGTTTATTCCTTGGTTTTCGTTGGTGATTGTCCCAGCTTTGTTTTTGGGTCTATTACTCATGG
TGGTATGGCCTAGTGTGGCCGCCGCTTACTGGCCTTGGGTGGATTGGACGTTTTTACCGCTCGATTGGGCTTTGCAGTTT
GCCGATGTAGGCTGGTGGGTGGTCCCCAGCAAAGTACAAGGTGTGGTCGCAGCGAGTGTGGCCATCCTCTTGCTTTATCG
ATTTATGAGCCTAAAAGCCTGCAGCTTATTATTGGGTATGATTGGCTTATGGTGGTGGTTTCCCTCTCTCACTCCACTTT
GGCGAATGGATGTGCTGGATGTTGGACATGGCTTGGCGATTGTGATTGAGCAAGATGAGCGAGCAATTGTCTACGATACA
GGCAGCAGTTGGCCGGGAGGCAGCTATGTGCAAAGCGTGATTGAGCCTATGCTCCAACAGCGGGGGCTACGCCAAGTGGA
TGGAGTGATTTTAAGTCATCTTGATAATGATCATGCGGGTGATTGGCAAGGTTTAGCTGAGCGCTGGCAACCCAATTGGA
TTCGTGCCAGCCAACTCGGGACAGAGTTTATGCCTTGTATCCGTGGTGAAAGCTGGCAGTGGCAATCTCTCCATTTTACG
GTGTTATGGCCACCACAAGCGGTTAGCCGAGCGTACAACCAGCATTCGTGTGTGATTCGTATGACCGATACTCAGTCTAA
CCATTCTGTACTGCTCTCCGGGGATGTCACAGCCATGGGGGAGTGGCTGCTTGCTCGCGACGGAGCGCAACTGCAAAGTG
AGGTGATGATCGTGCCGCACCACGGCAGTAAAACATCGTCCACCGCAGAGTTTATTGCCCAAGTGAATCCCAAACTTGCG
ATTGCTTCTGTGGCGAAAGATAACCGCTGGAATTTGCCTAATCCGCAAGTCGTGGCACGTTATCAAGCTCAGCAAGTTGA
GTGGCTAGATACTGGACACGCTGGGCAAATTAGCCTCTTTTTCTATCTAGATCAGCTGGATTGGTTTACCCAGCGTAGCC
TTGGCTGGCAGCCTTGGTATAGGCAGATGCTGCGTAAAGGAGTAGAATGA
3D structure
| Source | ID | Structure |
|---|
Similar proteins
Only experimentally validated proteins are listed.
| Protein | Organism | Identities (%) | Coverage (%) | Ha-value |
|---|---|---|---|---|
| comEC | Vibrio cholerae strain A1552 |
100 |
100 |
1 |
| comEC | Vibrio parahaemolyticus RIMD 2210633 |
41.433 |
100 |
0.416 |
| comEC | Vibrio campbellii strain DS40M4 |
41.301 |
99.718 |
0.412 |