Gene Daro_1164 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagDaro_1164 
Symbol 
ID3569290 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameDechloromonas aromatica RCB 
KingdomBacteria 
Replicon accessionNC_007298 
Strand
Start bp1268223 
End bp1269266 
Gene Length1044 bp 
Protein Length347 aa 
Translation table11 
GC content61% 
IMG OID637679631 
Productdihydroorotate oxidase A 
Protein accessionYP_284390 
Protein GI71906803 
COG category[F] Nucleotide transport and metabolism 
COG ID[COG0167] Dihydroorotate dehydrogenase 
TIGRFAM ID[TIGR01036] dihydroorotate dehydrogenase, subfamily 2 


Plasmid Coverage information

Num covering plasmid clones61 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones23 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGCTATATC CCCTGATCCG CAAATTCTTC TTCTCCCTCG ACGCTGAAAC CGCCCACGGT 
ATCGGCATGA AGGGCATCGA TCTGATGAAC GCCGCCGGCC TGGCCTGCGC CGTCGCCAAG
CCAGTCGCCG CCTGCCCGGT CGAAGTCATG GGCCTCAAAT TCCCCAATCC GGTCGGCCTG
GCCGCCGGTC TGGACAAGAA CGGCGATCAC ATCGATGGCC TGGCCAAGCT CGGCTTCGGC
TTCATCGAAA TCGGCACGAT CACGCCGCGC CCGCAGGATG GCAACCCGAA GCCGCGCCTG
TTCCGCATCC CGGAAGCGCA AGGCATCATC AACCGGATGG GCTTCAACAA CGCCGGCGTC
GACAAATTGC TGGAAAACGT CCGCGCCGCC GAATTCCCGA AAAAGGGCGG CATCCTCGGC
ATCAACATCG GCAAGAACGC GACGACGCCG ATCGAGAAGG CCGCCGACGA TTACCTGATC
TGCCTCGACA AGGTCTACAA CGACGCCAGC TACGTGACGG TCAATATCTC GTCGCCGAAC
ACCAAGAACC TGCGCGAATT GCAGAAGGAT GAGGCGCTCG ATGACCTGCT GGCACAACTG
AAAGCAAAAC AGCTGCAACT GGCCGAGCAA TACGGCAAGT ACGTGCCGAT GGCTCTGAAG
ATCGCTCCCG ATCTCGACGA CGAGCAGATC ACCGCCATCG CCGATGCGCT GCGCCGCCAT
CGCTTCGATG CGGTGATCGC CACCAACACC ACGCTGTCGC GCGAAGGCGT CGAAGGCATG
CCGAATGCCA CAGAAACAGG CGGCCTGTCG GGCAAGCCGG TCTTCGAGAA ATCAACAGCC
GTGCAGAAAA AACTATCGAT CGCCCTCGCC GGCGAACTTC CGATCATCGG TGTCGGCGGC
ATCATGGGCG GTGAGGATGC GGCCGAGAAA ATCCGCGCCG GCGCAAGCCT GGTGCAGTTC
TACAGCGGTT TCATCTACCG TGGCCCGGAT CTGGTCAGCG AGGTAGCAGA AACACTGGCC
CACGTCATGC GGAAATCCGT ATAA
 
Protein sequence
MLYPLIRKFF FSLDAETAHG IGMKGIDLMN AAGLACAVAK PVAACPVEVM GLKFPNPVGL 
AAGLDKNGDH IDGLAKLGFG FIEIGTITPR PQDGNPKPRL FRIPEAQGII NRMGFNNAGV
DKLLENVRAA EFPKKGGILG INIGKNATTP IEKAADDYLI CLDKVYNDAS YVTVNISSPN
TKNLRELQKD EALDDLLAQL KAKQLQLAEQ YGKYVPMALK IAPDLDDEQI TAIADALRRH
RFDAVIATNT TLSREGVEGM PNATETGGLS GKPVFEKSTA VQKKLSIALA GELPIIGVGG
IMGGEDAAEK IRAGASLVQF YSGFIYRGPD LVSEVAETLA HVMRKSV