Detailed information
Overview
| Name | comM | Type | Machinery gene |
| Locus tag | AW72_RS23715 | Genome accession | NZ_CP017975 |
| Coordinates | 4676719..4678239 (+) | Length | 506 a.a. |
| NCBI ID | WP_000050208.1 | Uniprot ID | - |
| Organism | Salmonella enterica subsp. enterica serovar Montevideo str. CDC 08-1942 | ||
| Function | require for natural transformation (predicted from homology) Unclear |
||
Genomic Context
Location: 4671719..4683239
| Locus tag | Gene name | Coordinates (strand) | Size (bp) | Protein ID | Product | Description |
|---|---|---|---|---|---|---|
| AW72_RS23695 (AW72_22515) | ilvE | 4673277..4674206 (-) | 930 | WP_000208532.1 | branched-chain-amino-acid transaminase | - |
| AW72_RS23700 (AW72_22520) | ilvM | 4674224..4674487 (-) | 264 | WP_000983260.1 | acetolactate synthase 2 small subunit | - |
| AW72_RS23705 (AW72_22525) | ilvG | 4674484..4676130 (-) | 1647 | WP_001012573.1 | acetolactate synthase 2 catalytic subunit | - |
| AW72_RS23975 | ilvX | 4676133..4676183 (-) | 51 | WP_192930503.1 | peptide IlvX | - |
| AW72_RS23710 (AW72_22530) | ilvL | 4676270..4676368 (-) | 99 | WP_001311244.1 | ilv operon leader peptide | - |
| AW72_RS23715 (AW72_22535) | comM | 4676719..4678239 (+) | 1521 | WP_000050208.1 | YifB family Mg chelatase-like AAA ATPase | Machinery gene |
| AW72_RS23720 (AW72_22540) | maoP | 4678264..4678602 (-) | 339 | WP_000840996.1 | macrodomain Ori organization protein MaoP | - |
| AW72_RS23725 (AW72_22545) | hdfR | 4678709..4679557 (+) | 849 | WP_001657826.1 | HTH-type transcriptional regulator HdfR | - |
Sequence
Protein
Download Length: 506 a.a. Molecular weight: 54957.18 Da Isoelectric Point: 8.4854
>NTDB_id=204757 AW72_RS23715 WP_000050208.1 4676719..4678239(+) (comM) [Salmonella enterica subsp. enterica serovar Montevideo str. CDC 08-1942]
MSLAIVHTRAALGVNAPPITIEVHISNGLPGLTMVGLPETTVKEARDRVRSAIINSGYEFPAKKITINLAPADLPKEGGR
YDLPIAVALLAASEQLTASNLEAYELVGELALTGALRGVPGAISSATEAIKAGRNIIVATENAAEVGLISKEGCFIADHL
QTVCAFLEGKHALERPLAQDMASPTATADLRDVIGQEQGKRGLEITAAGGHNLLLIGPPGTGKTMLASRLSGILPPLSNE
EALESAAILSLVNADTVQKRWQQRPFRSPHHSASLTAMVGGGAIPAPGEISLAHNGILFLDELPEFERRTLDALREPIES
GQIHLSRTRAKITYPAKFQLIAAMNPSPTGHYQGNHNRCTPEQTLRYLNRLSGPFLDRFDLSLEIPLPPPGILSQHASKG
ESSATVKKRVIAAHERQYRRQKKLNARLEGREIQKYCVLHHDDARWLEDTLVHLGLSIRAWQRLLKVARTIADIELADQI
SRQHLQEAVSYRAIDRLLIHLQKLLA
MSLAIVHTRAALGVNAPPITIEVHISNGLPGLTMVGLPETTVKEARDRVRSAIINSGYEFPAKKITINLAPADLPKEGGR
YDLPIAVALLAASEQLTASNLEAYELVGELALTGALRGVPGAISSATEAIKAGRNIIVATENAAEVGLISKEGCFIADHL
QTVCAFLEGKHALERPLAQDMASPTATADLRDVIGQEQGKRGLEITAAGGHNLLLIGPPGTGKTMLASRLSGILPPLSNE
EALESAAILSLVNADTVQKRWQQRPFRSPHHSASLTAMVGGGAIPAPGEISLAHNGILFLDELPEFERRTLDALREPIES
GQIHLSRTRAKITYPAKFQLIAAMNPSPTGHYQGNHNRCTPEQTLRYLNRLSGPFLDRFDLSLEIPLPPPGILSQHASKG
ESSATVKKRVIAAHERQYRRQKKLNARLEGREIQKYCVLHHDDARWLEDTLVHLGLSIRAWQRLLKVARTIADIELADQI
SRQHLQEAVSYRAIDRLLIHLQKLLA
Nucleotide
Download Length: 1521 bp
>NTDB_id=204757 AW72_RS23715 WP_000050208.1 4676719..4678239(+) (comM) [Salmonella enterica subsp. enterica serovar Montevideo str. CDC 08-1942]
ATGTCACTGGCGATTGTTCATACCCGCGCCGCACTTGGCGTCAATGCGCCGCCTATCACTATAGAGGTGCATATCAGTAA
TGGATTACCGGGATTAACCATGGTTGGCCTACCGGAAACCACGGTAAAAGAGGCGCGTGACCGGGTACGTAGTGCGATTA
TTAATAGCGGATATGAATTTCCGGCTAAAAAAATAACCATTAACCTTGCTCCAGCCGATCTACCGAAAGAGGGCGGAAGG
TATGACCTGCCTATTGCTGTTGCGCTTCTGGCCGCGTCTGAGCAGCTTACAGCGTCGAATCTTGAGGCATATGAGCTGGT
GGGTGAGTTAGCGCTTACAGGCGCATTACGCGGCGTTCCTGGCGCAATATCAAGTGCAACGGAAGCCATCAAGGCCGGCA
GAAATATTATCGTCGCAACAGAGAACGCGGCGGAGGTTGGGCTTATCAGCAAAGAAGGTTGTTTTATCGCCGATCATCTA
CAAACCGTCTGCGCCTTTCTGGAAGGGAAACACGCCCTGGAAAGACCTTTAGCTCAGGATATGGCATCGCCTACCGCAAC
TGCCGATCTTCGCGATGTGATCGGTCAGGAGCAGGGTAAACGCGGCCTGGAGATTACAGCGGCAGGGGGACATAATCTGC
TATTGATCGGCCCGCCGGGTACGGGTAAAACCATGCTGGCCAGTCGACTGAGCGGGATTCTTCCACCATTAAGCAATGAA
GAAGCATTGGAAAGCGCCGCGATCCTCAGTCTGGTTAATGCCGATACGGTACAAAAACGATGGCAGCAACGCCCCTTTCG
CTCACCTCATCATAGCGCCTCACTTACTGCTATGGTCGGCGGCGGCGCAATACCCGCCCCAGGAGAGATATCGCTGGCGC
ACAACGGAATTTTGTTCCTTGATGAATTGCCTGAATTTGAACGACGCACACTGGATGCGCTACGTGAACCTATAGAATCC
GGTCAAATCCATTTATCCCGTACCAGAGCGAAAATAACGTACCCTGCGAAGTTCCAGTTAATCGCCGCAATGAATCCCAG
CCCGACCGGACATTATCAGGGAAACCATAATCGCTGCACGCCAGAACAGACACTACGTTACCTTAATCGGTTGTCAGGCC
CGTTTCTTGATCGTTTTGACCTTTCGCTTGAGATACCGCTTCCACCGCCCGGGATTCTTAGCCAACATGCCTCAAAGGGT
GAGAGCAGCGCTACGGTAAAAAAGCGGGTCATCGCCGCCCATGAACGGCAGTACCGACGCCAGAAGAAGTTAAACGCGCG
TCTGGAGGGTCGCGAAATCCAAAAATATTGTGTATTGCATCACGATGACGCCCGCTGGCTTGAAGACACGCTGGTGCATC
TTGGATTATCCATTCGCGCCTGGCAGCGTTTACTAAAAGTGGCCAGAACCATTGCCGACATAGAACTGGCTGACCAGATC
TCGCGTCAGCATTTGCAGGAGGCGGTAAGCTATCGGGCGATAGACAGGTTGTTAATTCATTTGCAAAAGCTGTTGGCGTA
A
ATGTCACTGGCGATTGTTCATACCCGCGCCGCACTTGGCGTCAATGCGCCGCCTATCACTATAGAGGTGCATATCAGTAA
TGGATTACCGGGATTAACCATGGTTGGCCTACCGGAAACCACGGTAAAAGAGGCGCGTGACCGGGTACGTAGTGCGATTA
TTAATAGCGGATATGAATTTCCGGCTAAAAAAATAACCATTAACCTTGCTCCAGCCGATCTACCGAAAGAGGGCGGAAGG
TATGACCTGCCTATTGCTGTTGCGCTTCTGGCCGCGTCTGAGCAGCTTACAGCGTCGAATCTTGAGGCATATGAGCTGGT
GGGTGAGTTAGCGCTTACAGGCGCATTACGCGGCGTTCCTGGCGCAATATCAAGTGCAACGGAAGCCATCAAGGCCGGCA
GAAATATTATCGTCGCAACAGAGAACGCGGCGGAGGTTGGGCTTATCAGCAAAGAAGGTTGTTTTATCGCCGATCATCTA
CAAACCGTCTGCGCCTTTCTGGAAGGGAAACACGCCCTGGAAAGACCTTTAGCTCAGGATATGGCATCGCCTACCGCAAC
TGCCGATCTTCGCGATGTGATCGGTCAGGAGCAGGGTAAACGCGGCCTGGAGATTACAGCGGCAGGGGGACATAATCTGC
TATTGATCGGCCCGCCGGGTACGGGTAAAACCATGCTGGCCAGTCGACTGAGCGGGATTCTTCCACCATTAAGCAATGAA
GAAGCATTGGAAAGCGCCGCGATCCTCAGTCTGGTTAATGCCGATACGGTACAAAAACGATGGCAGCAACGCCCCTTTCG
CTCACCTCATCATAGCGCCTCACTTACTGCTATGGTCGGCGGCGGCGCAATACCCGCCCCAGGAGAGATATCGCTGGCGC
ACAACGGAATTTTGTTCCTTGATGAATTGCCTGAATTTGAACGACGCACACTGGATGCGCTACGTGAACCTATAGAATCC
GGTCAAATCCATTTATCCCGTACCAGAGCGAAAATAACGTACCCTGCGAAGTTCCAGTTAATCGCCGCAATGAATCCCAG
CCCGACCGGACATTATCAGGGAAACCATAATCGCTGCACGCCAGAACAGACACTACGTTACCTTAATCGGTTGTCAGGCC
CGTTTCTTGATCGTTTTGACCTTTCGCTTGAGATACCGCTTCCACCGCCCGGGATTCTTAGCCAACATGCCTCAAAGGGT
GAGAGCAGCGCTACGGTAAAAAAGCGGGTCATCGCCGCCCATGAACGGCAGTACCGACGCCAGAAGAAGTTAAACGCGCG
TCTGGAGGGTCGCGAAATCCAAAAATATTGTGTATTGCATCACGATGACGCCCGCTGGCTTGAAGACACGCTGGTGCATC
TTGGATTATCCATTCGCGCCTGGCAGCGTTTACTAAAAGTGGCCAGAACCATTGCCGACATAGAACTGGCTGACCAGATC
TCGCGTCAGCATTTGCAGGAGGCGGTAAGCTATCGGGCGATAGACAGGTTGTTAATTCATTTGCAAAAGCTGTTGGCGTA
A
3D structure
| Source | ID | Structure |
|---|
Similar proteins
Only experimentally validated proteins are listed.
| Protein | Organism | Identities (%) | Coverage (%) | Ha-value |
|---|---|---|---|---|
| comM | Haemophilus influenzae Rd KW20 |
58.317 |
100 |
0.589 |
| comM | Glaesserella parasuis strain SC1401 |
58.35 |
100 |
0.587 |
| comM | Vibrio cholerae strain A1552 |
58.614 |
99.802 |
0.585 |
| comM | Vibrio campbellii strain DS40M4 |
57.341 |
99.605 |
0.571 |
| comM | Legionella pneumophila str. Paris |
48.089 |
98.221 |
0.472 |
| comM | Legionella pneumophila strain ERS1305867 |
48.089 |
98.221 |
0.472 |
| RA0C_RS07335 | Riemerella anatipestifer ATCC 11845 = DSM 15868 |
42.998 |
100 |
0.431 |