Gene Dtox_4117 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagDtox_4117 
Symbol 
ID8431131 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameDesulfotomaculum acetoxidans DSM 771 
KingdomBacteria 
Replicon accessionNC_013216 
Strand
Start bp4286470 
End bp4287657 
Gene Length1188 bp 
Protein Length395 aa 
Translation table11 
GC content51% 
IMG OID645036312 
Productglycosyl transferase group 1 
Protein accessionYP_003193410 
Protein GI258517188 
COG category[M] Cell wall/membrane/envelope biogenesis 
COG ID[COG0438] Glycosyltransferase 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones11 
Plasmid unclonability p-value0.0723713 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones
Fosmid unclonability p-value0.0000000981139 
Fosmid HitchhikerYes 
Fosmid clonabilityhitchhiker 
 

Sequence

Gene sequence
GTGATGAAAC AGTTCCGGGC TGTGGTTTTG GTGGGAATGT ACGAGTGGCA TGACCGGGAG 
ATCAATGATA CGGTTAGGTG TCTGGCGCGT GCTTTCGAGG GTTTTGACCG CTACTATATC
GACCCGCCCA AAGGTCTGAG AGCCTTGAAA GGCAATATGT CCTATGTTTT AAAATCCCCG
CGCTGGCAAT GGTCACAGGA CGGTGAGGTG GCCGTGGGTG TCCCGCCTCT GGGTTTTCTG
CCGGTAAAGC TCGGTCTGAG GGAGCGCGCC AACCGCTGCG CGGCCTGCGG CCTAATTCGC
AGGCTGAAAA GAAATTATGG CGCTGATTGG CGCGAGCATA CTCTGTTCTA TGTTTCTTCC
GGCAGCTATA CGATAACCGG TTTTATTGAT ATGCTGGCTC CCAAGCACAT GGTCTTTCAC
CTGCTGGATG ATAACTTTGC CTTTCCCATT ATTAAGAATG ACCGCAGGGT TTGGGAAGAA
AATAAAGCAT TCATGGATTT TATGCTGCTG CACAGTTCAC TGGTGCTGGC GGTTTCACAG
GAATTGGTGG GCAAATACAG TGAAATGTAT AACAGAAAAA TATATCTGCT GGGCAACGGG
GTTGATGTGG AGCACTTCAG TCCGGAAAAC AAGGCCTGGC CGGAGGCGCC GGAACTGGCG
GGAATTTCAG AACCCGTGCT CCTGTTTATC GGAGCGGTTA ATTCCTGGAT TGATATAGGG
CTGCTTAAGG AGTTGGCCGA AAAGAGACCT GCTTACAAGT TGGTGATAAT CGGCCCCTGT
TACGAAAGCA GCATTGATTT GGCTGTCTGG AACAGCCTGA AGGAAATGAG CGGCGTGCTG
TGGCTGGGCA GCAGGCCTTT TGCCGAGCTG CCTCACTATA TTCAACATGC TTCGGTACTC
CTGCTGCCCA GAACCAGGGA TGAACACTCG TTGGCCTCCG ACCCGCTGAA ACTTTACGAG
TACCTGGCTA CCGGTAAGCC TGTGGTGGCG GTAGGGATTC CGGCGGTACA AAAATTTGCC
GCTTTTGTTT ATGCGGCTGC CGGCAGAGAA GATTTTATTA AGCTTACAGA TCGGGCATTG
TCCGAGTGGA ATGAGGAAAA ACAGTCATTA CAGCTGGCTG CGGCTGAGGA ATATTCCTGG
TCTTCCAGAA TTGGCACAAT TCTAAAATTG CTGGCCGAGA CCGGATAA
 
Protein sequence
MMKQFRAVVL VGMYEWHDRE INDTVRCLAR AFEGFDRYYI DPPKGLRALK GNMSYVLKSP 
RWQWSQDGEV AVGVPPLGFL PVKLGLRERA NRCAACGLIR RLKRNYGADW REHTLFYVSS
GSYTITGFID MLAPKHMVFH LLDDNFAFPI IKNDRRVWEE NKAFMDFMLL HSSLVLAVSQ
ELVGKYSEMY NRKIYLLGNG VDVEHFSPEN KAWPEAPELA GISEPVLLFI GAVNSWIDIG
LLKELAEKRP AYKLVIIGPC YESSIDLAVW NSLKEMSGVL WLGSRPFAEL PHYIQHASVL
LLPRTRDEHS LASDPLKLYE YLATGKPVVA VGIPAVQKFA AFVYAAAGRE DFIKLTDRAL
SEWNEEKQSL QLAAAEEYSW SSRIGTILKL LAETG