Gene Dtox_3441 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagDtox_3441 
Symbol 
ID8430435 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameDesulfotomaculum acetoxidans DSM 771 
KingdomBacteria 
Replicon accessionNC_013216 
Strand
Start bp3641848 
End bp3642870 
Gene Length1023 bp 
Protein Length340 aa 
Translation table11 
GC content46% 
IMG OID645035668 
Productmolybdopterin binding domain protein 
Protein accessionYP_003192787 
Protein GI258516565 
COG category[H] Coenzyme transport and metabolism 
COG ID[COG0303] Molybdopterin biosynthesis enzyme 
TIGRFAM ID[TIGR00177] molybdenum cofactor synthesis domain 


Plasmid Coverage information

Num covering plasmid clones17 
Plasmid unclonability p-value0.593613 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones32 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGAAAGTTG TAGCAGTAGA AGAGGCTGTG GGCATGGTGT TGGGTCATGA TATTACGCAA 
ATAGTTCCCG GTAAAGTAAA GGGCCCTGCT TTTAAGAAAG GTCATGTAAT TGAATTTAAA
GATATACCGA AACTTTTAAA TATCGGGAAA GAAAATATCT ATGTATTTGA TTTGCAGGAG
GGCTTTGTTC ATGAAGATGA TGCCGCTCTT AGGATTGCAG AGGCGGCTGC CGGTCCCGGC
ATTGAGATAT CCCAGCCTAA AGAGGGACGT GTTAACCTTA CGGCAACTGT TTCCGGCTTG
TTGAAAATTA ATGCCGGAGC CTTATTCCGC ATTAATGAAA TTGATGAAGC GGCCATGGCC
ACTTTGCATA CAAACCACAG AATAGATGCC GGCAGCGTTG TTGCCGGTAC CAGAATAATA
CCGCTGGTTA TTGAGGAAGA AAAGCTTGCT TCCATAATTA ACGTGTGCAG AGAAAATTAC
CCGATAGCGG AGGTTAAGCC GTTTAAACCC TGGAAGATAG GTGTTGTGAC AACCGGCAGT
GAGGTTTACC GCGGCAGGAT CAAAGACAAG TTCGGCCCGG TATTGAAATC CAAATTTTCT
CTGTTAGGCA GCACTGTTTT AAGGCAGATC TTCGTGTCTG ATGATGTACA GATGACAGCT
CAGGCGGTAC ATGATTTGAT CGATGAAGGG GCTGATATGA TTGCCGTTAC CGGCGGTATG
TCGGTAGATC CTGATGATCA GACTCCGGCC GGTATCCGTG CTGCCGGTGG GGAAGTTGTT
ATATACGGTA CTCCTTGTTT GCCCGGCTCC ATGTTTATGC TGGCATATAT TAATGATACT
CCTGTTGTGG GCCTGCCGGG TTGTGTTATG TATAACAAGG CGAGTATTTT TGATTTAGTT
GTGCCGCGCA TACTGGCCGG TGAAGAAGTG ACGAAAGAGG ATATTATAGC TCTTGGTCAT
GGAGGCTTAT GCTCGGGTTG TCCTGAATGC AGGTACCCGG TTTGTGGTTT TGGCAAGGTG
TAA
 
Protein sequence
MKVVAVEEAV GMVLGHDITQ IVPGKVKGPA FKKGHVIEFK DIPKLLNIGK ENIYVFDLQE 
GFVHEDDAAL RIAEAAAGPG IEISQPKEGR VNLTATVSGL LKINAGALFR INEIDEAAMA
TLHTNHRIDA GSVVAGTRII PLVIEEEKLA SIINVCRENY PIAEVKPFKP WKIGVVTTGS
EVYRGRIKDK FGPVLKSKFS LLGSTVLRQI FVSDDVQMTA QAVHDLIDEG ADMIAVTGGM
SVDPDDQTPA GIRAAGGEVV IYGTPCLPGS MFMLAYINDT PVVGLPGCVM YNKASIFDLV
VPRILAGEEV TKEDIIALGH GGLCSGCPEC RYPVCGFGKV