Gene RPB_3652 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagRPB_3652 
Symbol 
ID3911454 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameRhodopseudomonas palustris HaA2 
KingdomBacteria 
Replicon accessionNC_007778 
Strand
Start bp4192015 
End bp4193235 
Gene Length1221 bp 
Protein Length406 aa 
Translation table11 
GC content66% 
IMG OID637885554 
Productcytochrome P450 
Protein accessionYP_487258 
Protein GI86750762 
COG category[Q] Secondary metabolites biosynthesis, transport and catabolism 
COG ID[COG2124] Cytochrome P450 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones20 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones11 
Fosmid unclonability p-value0.348066 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGGAAATGC CCTCACGCGA ACTGGCGGCG GAGTTCGAAC TCGAACGGCT GACGCCGGAG 
TTCTATGACA ATCCCTACCC CACCTATCGC GCCCTGCAGA CGCATCAGCC GGTCAAGCGG
CTGCGCAATG GCGGGTACAT CCTGACGCGC TATGACGATC TGGTGACGGT CTACAAGAAC
ACCACGCTGT TCAGCTCCGA CAAGAAGCGC GAATTCGCGC CGAAATACGG CGACTCGCTG
CTGTTCGAGC ATCACACTAC CAGCCTGGTG TTCAACGACC CGCCGGCGCA TACGCGAGTG
CGGCGGCTGA TCACGGGCGC GCTGTCGCCG CGCGCGATCG CCGGCATGCA GCCGGATCTG
ATCGCGCTGG TCGACCGCCT GCTCGATGCG ATGGCGGCCA AAGCCGGCGT CGATCTGATC
GAGGATTTCG CCGCCGCGAT CCCGATCGAG GTGATCGGCA ATCTGCTCGG CGTGCCGCAC
GACGAGCGCG GCCCGCTCCG CGACTGGTCG CTGGCGATTC TCGGCGCACT CGAGCCGGTG
ATCGGGCCGG AGACGTTTTC GCGCGGCAAT GAGGCTGTCC GCGACTTCCT CGCCTATCTC
GAAATCCTGA TCACGCGCCG TCGCGCCGAG CCCGGCGATC CGGAGCACGA TGTTCTGACC
CGGCTGATCC AGGGCGACGA CGGCACCGGC GAGAAGCTCT CCGCCAAGGA GCTGCTGCAC
AATTGCATCT TCCTGCTCAA CGCCGGACAT GAAACCACCA CCAACCTGAT CGGCAACGGG
CTCGTGGCGC TCGCAGACAA TCCTGCGGAA AAACAGCGGC TGATCGGCCA GCCCGGCCTC
GCCCGCACCG CGGTCGAAGA GATCCTGCGC TATGAGAGCT CGAACCAGCT CGGCAACCGC
ATCACCACCA CCGAGGTCGA GATCGGAGGC GTAACGATGC AGGCCAACAC CTCGCTGACG
CTGTGCATCG GCGCTGCCAA CCGCGATCCG GCGCAGTTTC CCGATCCCGA CCGGTTCGAC
GTCGGACGAA CGCCGAACCG GCACCTCGCT TTTGCCACGG GGCCACATCA ATGCGCCGGC
ATGGCGCTGG CGCGGCTCGA AGGCGTGATC GCGCTGACGC GATTCCTGGC GCGCTTCCCG
AACTACACGC TCGACGGCAC GCCGTCGCGC GGCGGGCGGG TGCGGTTTCG CGGCTATCTG
CGCGTGCCAT GCCGCCTGTA G
 
Protein sequence
MEMPSRELAA EFELERLTPE FYDNPYPTYR ALQTHQPVKR LRNGGYILTR YDDLVTVYKN 
TTLFSSDKKR EFAPKYGDSL LFEHHTTSLV FNDPPAHTRV RRLITGALSP RAIAGMQPDL
IALVDRLLDA MAAKAGVDLI EDFAAAIPIE VIGNLLGVPH DERGPLRDWS LAILGALEPV
IGPETFSRGN EAVRDFLAYL EILITRRRAE PGDPEHDVLT RLIQGDDGTG EKLSAKELLH
NCIFLLNAGH ETTTNLIGNG LVALADNPAE KQRLIGQPGL ARTAVEEILR YESSNQLGNR
ITTTEVEIGG VTMQANTSLT LCIGAANRDP AQFPDPDRFD VGRTPNRHLA FATGPHQCAG
MALARLEGVI ALTRFLARFP NYTLDGTPSR GGRVRFRGYL RVPCRL