Gene Acel_1296 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagAcel_1296 
Symbol 
ID4485637 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameAcidothermus cellulolyticus 11B 
KingdomBacteria 
Replicon accessionNC_008578 
Strand
Start bp1444355 
End bp1445410 
Gene Length1056 bp 
Protein Length351 aa 
Translation table11 
GC content70% 
IMG OID639730076 
Productdihydroorotate oxidase B, catalytic subunit 
Protein accessionYP_873054 
Protein GI117928503 
COG category[F] Nucleotide transport and metabolism 
COG ID[COG0167] Dihydroorotate dehydrogenase 
TIGRFAM ID[TIGR01037] dihydroorotate dehydrogenase (subfamily 1) family protein 


Plasmid Coverage information

Num covering plasmid clones19 
Plasmid unclonability p-value0.717011 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones23 
Fosmid unclonability p-value0.119624 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGACGACCA CCGGTTCACC GCACACCGGC TTTGCGGTGC GGGCAACACC ACGCCGGACG 
GCCGCACCGG CTGTCGATCT GACGACGCAG CTGGGATCCG TGGTCCTGCC CAATCCGGTC
ACGACGGCGT CCGGTTGCGC CGCGGCAGGG CGCGAGCTCG GCCAGTTCGT CGACGTTTCC
ACGCTGGGTG CGGTCGTGAC AAAATCGATC ATGCTGGCGC CTCGGGCGGG ACGGCCGACA
CCGCGGATGG CGGAAACGCC GAGTGGACTG CTCAACGGCA TCGGCCTGCA AGGTCCGGGA
ATCGACGAAT TCCTTGAGCA TGATCTGCCG TGGCTGGCTG AGCATGGCGC CCGGGCGATC
GTGTCGATTG CCGGATCGAG CGTCGGGGAG TACGCGGCGC TCGCCGCCCG GCTCCGGGGT
GCCGCTCCGG TCGTGGCGCT CGAGGTGAAC ATCTCCTGCC CCAACGTGGA GGACCGCGGC
CAGGTGTTCG CCTGCGATCC GCGGGCCGCC AGCGCTGTGC TGGCCGCCGT CCGGGAAGCC
GCGGACCCGG CGGTCCCTGT CTTCGCGAAA CTCTCCCCGG ACGTCACCGA CATCGTGGCC
GTCGCCACGG CGTGTGTCGC CGCCGGCGCG GACGGCCTCT CGCTGATCAA CACGCTCCTC
GGGCTGGTGA TCGATACCGA AACCCTGCGG CCCGCGCTCT CCGGGGTCAC CGGTGGTCTG
TCAGGTCCGG CGATCCGCCC GGTCGCGGTC CGCTGTGTCT GGCAGGTGCA CGCCGCACTC
CCGGACGTAC CGATCATCGG CATGGGTGGA ATACGAAGCG GCTTGGATGC CCTGCAGTTC
CTGCTCGCCG GAGCGTGTGC GGTCAGCGTC GGGACGGAAA TCTTCCACGA TCCGAGCGCA
CCAGCCCGGA TCCGTGATGA GCTTGCCGAG GCGCTTGCCG CCCGGGGCTT CCAGCGGGTC
AGCGATGTGA TCGGCCTGGC GCACCGCGGC GGTTCCGTGG CGGAGGGTGG CGATTCCCTG
GCGCGCGGTG ACGATTTCCT CGGCGCTCGG CGATAG
 
Protein sequence
MTTTGSPHTG FAVRATPRRT AAPAVDLTTQ LGSVVLPNPV TTASGCAAAG RELGQFVDVS 
TLGAVVTKSI MLAPRAGRPT PRMAETPSGL LNGIGLQGPG IDEFLEHDLP WLAEHGARAI
VSIAGSSVGE YAALAARLRG AAPVVALEVN ISCPNVEDRG QVFACDPRAA SAVLAAVREA
ADPAVPVFAK LSPDVTDIVA VATACVAAGA DGLSLINTLL GLVIDTETLR PALSGVTGGL
SGPAIRPVAV RCVWQVHAAL PDVPIIGMGG IRSGLDALQF LLAGACAVSV GTEIFHDPSA
PARIRDELAE ALAARGFQRV SDVIGLAHRG GSVAEGGDSL ARGDDFLGAR R