Gene Saro_3172 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagSaro_3172 
Symbol 
ID3918214 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameNovosphingobium aromaticivorans DSM 12444 
KingdomBacteria 
Replicon accessionNC_007794 
Strand
Start bp3387136 
End bp3388131 
Gene Length996 bp 
Protein Length331 aa 
Translation table11 
GC content65% 
IMG OID640445956 
Productaldo/keto reductase 
Protein accessionYP_498441 
Protein GI87201184 
COG category[C] Energy production and conversion 
COG ID[COG0667] Predicted oxidoreductases (related to aryl-alcohol dehydrogenases) 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones30 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGTCGATTC AGCGCCGCCC GATCAACGGA CGCGAGACCA ACCCCATTGG TCTGGGATGC 
ATGTCGCTGA GCTGGGCCTA TGGCGGCCGG CCGAGCGACG AGGACGGCAT CCGCCTGCTC
CAGCACGCCG TCGACATCGG CTACGATCAT TTCGACACCG CGCGGCTCTA TGGTCTCGGC
CACAACGAGA CCCAGGTCGG GATCGCGCTC AAGGGCCGAC GCGACAAGGT TTTCCTGGCC
TCGAAGATGG GCATCTTCGC CAGCGGCGAC AAGCGCGGCA TCGATTGCCA CCCGGACACG
ATCCGCAGCG AACTCGAAGT CTCGCTCAGG CTGCTCCAGA CCGACCACAT CGACCTCTAC
TACATGCACC GCCGCGATTT CACCGTGCCG ATCGAGGATT CGGTCGGCGC GATGGCCGAC
CTCGTGAAGG AGGGCAAGAT CGGCGGCATC GGCCTGTCCG AAATGTCGGC TGACACGCTG
CGCAAGGCTT CGGCGGTCCA CCCCATCGCC GCGATGCAGA CCGAATATTC ACCCTGGACC
CGCCAGGCCG AAATCGCCGT CCTCGAGGCC TGCCGCGAGC TTGGCACCAC GTTCGTCGCC
TTTTCGCCGG TCGCGCGCGG GGTTCTGGCC GATGGCGTGC ACGATCCCGC CGCGCTCGAG
GAAAAGGACA TCCGGCGCGC CATGCCGCGC TTCATGGGCG ACAACTGGCC CAGGAACTAC
GCGCTCGTCC GCCAGTTCAA TGCCATCGCC GCTCGCGAAG GCGTGACCCC GGCGCAGCTT
TCGCTCGCCT GGGTCCTGTC GCGGGGCGAA CACGTCGTTG CCATTCCCGG CACCGGCAAG
ATCGCTCACC TCGAAGAGAA CATCGCACGC TGGGACTGGG AAATCCCGGT TGCGGTCGCT
GCCGAAGTCG ATGCCCTGAT CAACCAGCAG ACCGTCGCCG GTCACCGCTA TGCCGGGGTC
ATGCTGCCGA CGATCGATAC CGAGGATTTC GACTGA
 
Protein sequence
MSIQRRPING RETNPIGLGC MSLSWAYGGR PSDEDGIRLL QHAVDIGYDH FDTARLYGLG 
HNETQVGIAL KGRRDKVFLA SKMGIFASGD KRGIDCHPDT IRSELEVSLR LLQTDHIDLY
YMHRRDFTVP IEDSVGAMAD LVKEGKIGGI GLSEMSADTL RKASAVHPIA AMQTEYSPWT
RQAEIAVLEA CRELGTTFVA FSPVARGVLA DGVHDPAALE EKDIRRAMPR FMGDNWPRNY
ALVRQFNAIA AREGVTPAQL SLAWVLSRGE HVVAIPGTGK IAHLEENIAR WDWEIPVAVA
AEVDALINQQ TVAGHRYAGV MLPTIDTEDF D