Gene Moth_1446 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagMoth_1446 
Symbol 
ID3832615 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameMoorella thermoacetica ATCC 39073 
KingdomBacteria 
Replicon accessionNC_007644 
Strand
Start bp1488131 
End bp1489276 
Gene Length1146 bp 
Protein Length381 aa 
Translation table11 
GC content49% 
IMG OID637829379 
Product2-hydroxyglutaryl-CoA dehydratase, D-component 
Protein accessionYP_430299 
Protein GI83590290 
COG category[E] Amino acid transport and metabolism 
COG ID[COG1775] Benzoyl-CoA reductase/2-hydroxyglutaryl-CoA dehydratase subunit, BcrC/BadD/HgdB 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones23 
Plasmid unclonability p-value0.629308 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones11 
Fosmid unclonability p-value0.054659 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGGCCAATC AACCACTGGC GTATTTTGAT GATTTGCGGG AAAGAAATGT CCTGGAAATC 
AAAAAGTTGA AAGACCAGGG GAAAAAGGTA GTTGGCACTT ATTGTGCCTT TACCCCCAAG
GAACTCATCA TTGCCGCCGG TGCCATCCCA GTTTCCTTAT GTGGTACCAG GCAAGAACCG
ATTCCCGAGG CTGAAAAAAT CCTTCCCCGG AACCTCTGTC CGTTAATTAA ATCCAGTTTT
GGCTTTGCCA TTACCGGGAG TTGTCCTTAT TTTTATTTTG CCGACCTGCT TATTGCCGAA
ACCACCTGTG ATGGCAAGAT CAAGATGTAC GAATTGCTCA GGGAATATAA ACCCATGCAT
ATTCTCAATT TACCGCCCAC CTCGCTGGGC GAAGACGCCT TTGCCTACTG GTATAATGAA
ATACTTAAAG CCAAAGAACG GCTGGAAAGG GAATTTGCTG TTGAAATCAC AACGGCAAAA
TTGCAGGAAG CCATTCGCCT GGTTAATGAA GAACGCCGGG CTTTGCTGGA ATTTCATCGC
TTAAACCGGC ACGACCCAGC ACCGTTATCG GGTTTGGATC TTTTGAAGGT GCTTTGGGCA
AAGGGCTTTA CTCCCGATAT AGCCGCCGGT ACGGCCGTGA TTCGCCAGGT CACTCTGGCG
GTAAAGGAAC AGATGGCCAG GGGCGTGTCT GCGGCACCAC CGGGCTCACC TCGAATACTC
TTAACCGGTT GTCCGGTAGG CCTGGGTTCC GAGAAGGTGA TTAAACTCGT GGAGGCCGGG
GGCGGGGTGG TTGTTTGCCT GGAATCCTGC AGCGGCATTA AAGCCCTGGA GCCCCTTGTG
GATGAGGAAG GCGATCCCTT GCAGGCAATT GCTGCCAAGT ATTTGCAAGT ACCCTGTCCC
TGTTTAACCC CCAATCGGGG CCGTCTAGAA TTGCTGGAAC GTCTTATTAA AGAATACAGG
GTCGATGGGG TAATCGATCT CACCTGGCAG GCCTGCCATA CGTATAATAT TGAGTCTTAT
AGCATTAAGA AATTGGTCCA GGAAAAAGAG GGACTGCCTT TTCTACCAAT TGAGACTGAC
TATTCAACAA GTGACTTGCA GCAGCTCAAG GTCCGGATTG ACGCTTTCCT GGAAATGATT
AAATAG
 
Protein sequence
MANQPLAYFD DLRERNVLEI KKLKDQGKKV VGTYCAFTPK ELIIAAGAIP VSLCGTRQEP 
IPEAEKILPR NLCPLIKSSF GFAITGSCPY FYFADLLIAE TTCDGKIKMY ELLREYKPMH
ILNLPPTSLG EDAFAYWYNE ILKAKERLER EFAVEITTAK LQEAIRLVNE ERRALLEFHR
LNRHDPAPLS GLDLLKVLWA KGFTPDIAAG TAVIRQVTLA VKEQMARGVS AAPPGSPRIL
LTGCPVGLGS EKVIKLVEAG GGVVVCLESC SGIKALEPLV DEEGDPLQAI AAKYLQVPCP
CLTPNRGRLE LLERLIKEYR VDGVIDLTWQ ACHTYNIESY SIKKLVQEKE GLPFLPIETD
YSTSDLQQLK VRIDAFLEMI K