Gene RPD_2020 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagRPD_2020 
Symbol 
ID4022502 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameRhodopseudomonas palustris BisB5 
KingdomBacteria 
Replicon accessionNC_007958 
Strand
Start bp2265961 
End bp2266929 
Gene Length969 bp 
Protein Length322 aa 
Translation table11 
GC content68% 
IMG OID637962213 
Productzinc-binding alcohol dehydrogenase 
Protein accessionYP_569156 
Protein GI91976497 
COG category[C] Energy production and conversion
[R] General function prediction only 
COG ID[COG0604] NADPH:quinone reductase and related Zn-dependent oxidoreductases 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones15 
Plasmid unclonability p-value0.322186 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones13 
Fosmid unclonability p-value0.588155 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGAAGGCCT ATGTCTATGG CGCCGGCGGC GCCGCCATTA CCGACGTGGA CAAGCCGGCC 
CCGAAGGGAC CGCAGGTCCT GATCCGCGTC CGCGCCTGCG GCCTCAATCG CGCCGACCTC
GGCATGACCA AGGGCCATGC CCACGGCGCG GCCGGCGGCG TCGGCGCCGT GCTCGGCATG
GAATGGGCCG GTGAGATCGC AGCGGTCGGC GACGAGGCCT ATGGCTGCGA GGTCGGCGAC
CGCGTGATGG GTTCGGGCGC CGCGGCGTTC GGCGAGTACA CGCTGGCCGA TCACGGCCGG
CTGTTCCCGA TACCGGGCGG CATGAGCTTC GAGGATGCCG CCGCGCTGCC CGTCGCGCTC
ACCACCATGC ACAACGCGCT GATCGCGGTC GGCAAGCTGC GCGCCGGCCA GTCCGTGCTG
ATCCAGGGCG CGAGCTCCGG TGTCGGGCTG ATGGCGCTGC AGATCGCCCG GCTGAAGGGC
GCCAAGCTGG TGATCGGCTC GTCGACCGAC AACAGCCGCC GCGACCGGCT GCGCGAGTTC
GGCGCGGACC TCGCGATCGA TTCGTCGGGA AGCGGCTGGG TCGATCAGGT GCTCGCCGCC
ACCGGCGGCG CCGGGGTCGA TCTGATCATC GACCAGATTT CCGGCAGCGT CGCCAACCAG
AACCTCGCCG CGACCCGCGT GCTCGGCCGC ATCGTCAATG TCGGCCGGCT CGGCGGCGCC
CACGCCGATT TCAATTTCGA TCTGCATGCG GCGCGGCGGA TCGACTATGT CGGCGTCACC
TTCCGCACCC GCAGCATTGA AGAGATTCGC GAAATCTTCC GCCAGGTGCG CGGCGATATC
TGGCCGGCTG TCGAAACACG CAAGCTGAAG TTGCCGGTGG ACCGCGTCTT CCCGTTCGCC
GAGATCGACA AGGCCTTCGC ACACATGGAG GCAAACCGTC ATTTCGGAAA GATCGTCGTA
ACGCTCTGA
 
Protein sequence
MKAYVYGAGG AAITDVDKPA PKGPQVLIRV RACGLNRADL GMTKGHAHGA AGGVGAVLGM 
EWAGEIAAVG DEAYGCEVGD RVMGSGAAAF GEYTLADHGR LFPIPGGMSF EDAAALPVAL
TTMHNALIAV GKLRAGQSVL IQGASSGVGL MALQIARLKG AKLVIGSSTD NSRRDRLREF
GADLAIDSSG SGWVDQVLAA TGGAGVDLII DQISGSVANQ NLAATRVLGR IVNVGRLGGA
HADFNFDLHA ARRIDYVGVT FRTRSIEEIR EIFRQVRGDI WPAVETRKLK LPVDRVFPFA
EIDKAFAHME ANRHFGKIVV TL