Gene Dtox_3202 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagDtox_3202 
Symbol 
ID8430196 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameDesulfotomaculum acetoxidans DSM 771 
KingdomBacteria 
Replicon accessionNC_013216 
Strand
Start bp3403404 
End bp3404465 
Gene Length1062 bp 
Protein Length353 aa 
Translation table11 
GC content52% 
IMG OID645035448 
Product1-hydroxy-2-methyl-2-(E)-butenyl 4-diphosphate synthase 
Protein accessionYP_003192567 
Protein GI258516345 
COG category[I] Lipid transport and metabolism 
COG ID[COG0821] Enzyme involved in the deoxyxylulose pathway of isoprenoid biosynthesis 
TIGRFAM ID[TIGR00612] 1-hydroxy-2-methyl-2-(E)-butenyl 4-diphosphate synthase 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.00281506 
Plasmid hitchhikingYes 
Plasmid clonabilityhitchhiker 
 

Fosmid Coverage information

Num covering fosmid clones
Fosmid unclonability p-value0.00000333931 
Fosmid HitchhikerYes 
Fosmid clonabilityhitchhiker 
 

Sequence

Gene sequence
ATGAGGCGAA GGAAAACCAG GGTTATCCGG GTTGGTCAGG TGGCAGTCGG CGGTGACGCG 
CCGGTCAGTG TGCAGTCCAT GACCAACACG GATACCAGAG ACGCTGCGGC AACTGCCAGG
CAGATCAGGG AGCTGGCTCT GGCCGGGTGT GAAATTGTGC GGGTGGCAGT GCCGGATGAG
CAGGCGGCGG TCGCTTTAAA GGAAATCAAG AACGGTATTA ACGTGCCCTT GATTGCGGAT
ATTCATTTTG ATTATAAGCT GGCCCTGCAG GCTATCCGGG CAGGTGTAGA CGGACTGCGC
ATCAATCCGG GCAATATTGG CAGCAGGTCT AAAGTGGCCG AGGTGGTCAG GGCTGCCGGG
GACGCTCAGG TTCCTATCAG AATCGGGGTC AATGCCGGTT CATTGGAAAA AGAGCTGTTG
GATAAGTATG GAAATATAAC AGCTCAGGCT ATGGTGGAAA GCGCTCTCGG TCATATCGGC
ATTTTAGAAG ATTTAAATTT CAGTGATATA AAGATCTCTT TAAAGGCCTC AGACATTCCT
TTAATGCTGG AGGCCTATCG CCTTTTGGCC GACCGGGTGG ACTATCCTTT TCATATAGGT
GTAACCGAGG CGGGTACTTT GCGCTCCGGT ATTGTCAAGT CAGCTGTGGG CATCGGGGCA
CTGCTGGCTG AAGGTATCGG TGATACGCTC AGGGTTTCAC TGACAGGGGA CCCACTGCAC
GAGGTAAAGA CTGGCTATGA GATTCTAAAG GCTTTGGGCC TGCGCCAGAG AGGCATTGAA
TTTATTTCCT GTCCCACCTG CGGGCGGACT CAGATTGATT TAATTCGTAT TGCCAATGAG
GTGGAGGACA GGCTGCAGTT TGTAGATAAG CCGCTCAAGG TGGCGGTGAT GGGCTGTTCG
GTCAACGGTC CCGGTGAGGC GCGGGAGGCC GATATTGGGA TTGCCGGTGG CAGGGGGGCC
GGGTTATTGT TTAAAAAAGG CCGGACTGTG CGAAAGATTG AGGAAGCGGA TTTGGTGGAA
GAGCTTATTA AAGAAGTAGA AAATATGTTG GAAAATTATT AA
 
Protein sequence
MRRRKTRVIR VGQVAVGGDA PVSVQSMTNT DTRDAAATAR QIRELALAGC EIVRVAVPDE 
QAAVALKEIK NGINVPLIAD IHFDYKLALQ AIRAGVDGLR INPGNIGSRS KVAEVVRAAG
DAQVPIRIGV NAGSLEKELL DKYGNITAQA MVESALGHIG ILEDLNFSDI KISLKASDIP
LMLEAYRLLA DRVDYPFHIG VTEAGTLRSG IVKSAVGIGA LLAEGIGDTL RVSLTGDPLH
EVKTGYEILK ALGLRQRGIE FISCPTCGRT QIDLIRIANE VEDRLQFVDK PLKVAVMGCS
VNGPGEAREA DIGIAGGRGA GLLFKKGRTV RKIEEADLVE ELIKEVENML ENY