Gene Pnap_0117 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagPnap_0117 
Symbol 
ID4688223 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism namePolaromonas naphthalenivorans CJ2 
KingdomBacteria 
Replicon accessionNC_008781 
Strand
Start bp126045 
End bp128864 
Gene Length2820 bp 
Protein Length939 aa 
Translation table11 
GC content63% 
IMG OID639833110 
Productmethionine synthase 
Protein accessionYP_980363 
Protein GI121603034 
COG category[E] Amino acid transport and metabolism 
COG ID[COG1410] Methionine synthase I, cobalamin-binding domain 
TIGRFAM ID[TIGR00640] methylmalonyl-CoA mutase C-terminal domain
[TIGR02082] 5-methyltetrahydrofolate--homocysteine methyltransferase 


Plasmid Coverage information

Num covering plasmid clones15 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones13 
Fosmid unclonability p-value0.0603491 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGTCCTCCA TTCAGTTTTC CCATGAAGTC CCGCCCATGA AGCTGTCCGG CCTGGAGCCG 
GTCACCATCG GCAGCGATTC GCTGTTCGTC AACATCGGCG AGCGCACCAA CGTCACCGGC
TCCAAGGCCT TCGCCCGCAT GATTTTGAAC GGCGACTATG AACAGGCGCT GACAGTGGCG
CGCCAGCAGG TCGAAAACGG CGCGCAGATC ATCGACATCA ACATGGACGA GGCCATGCTC
GACAGCCAGG CGGCCATGGT GCGCTTTTTG AACCTGATCG CCGGCGAGCC CGACATTGCG
CGCGTTCCCA TCATGATCGA CAGCTCCAAG TGGAGCGTGA TCGAGGCCGG CCTGCGCTGC
ATCCAGGGCA AGGGCATCGT CAATTCGATT TCGATGAAGG AGGGCGTTGA CGCCTTCAAG
CACCAGGCCA AACTGCTCAA GCGCTACGGC GCCGCCGCCG TTGTCATGGC CTTCGATGAA
AAAGGCCAGG CCGACACCTA CGAGCGAAAA ATCAGCATCT GCGAGCGCGC TTACCGCGTG
CTGGTCGATG AGATCGGTTT TCCGCCCGAA GACATCATTT TTGACCCCAA TATCTTCGCC
ATCGCCACCG GCATCGACGA GCACAACAAC TACGCGGTCG ATTTCATCGA GGCCACGCGC
TGGATCAAGG CCAATTTGCC GGGCGCCAAG GTGTCGGGCG GCGTGTCCAA CGTGTCGTTC
AGCTTTCGCG GCAACGACCC GGTGCGCGAA GCGATTCACA CGGTGTTTTT GTACCACGCG
ATCCAGGCCG GCATGGACAT GGGCATCGTC AACGCCGGCA TGATGGGCGT CTATGACGAG
CTGGAGCCGG TGCTGCGCGA GCGCGTCGAG GACGTGGTGC TGAACCGCCA GCCGGTGTAC
AAGCCGGGCG AAGCGCATCT CACGCCCGGC GAACGGCTGA TCGAAGTCGC CGAAACGGCC
AAAAGCGGCG CGCGTGACGA CAGCAAGAAG TACGAATGGC GCGCGCTGCC GATACGCGCG
CGCCTGTCGC ATGCGCTGGT GCATGGCAAC AACGAATTCA TCACCGAGGA CACCGAGGAA
GTCTGGCAGG CCATCAAGGC CGAGGGTGGC CGCCCGCTGC ACGTCATCGA AGGCCCGCTG
ATGGATGGCA TGAACGTCGT CGGCGACCTG TTCGGCCAGG GCAAGATGTT TTTGCCGCAG
GTGGTGAAAA GCGCGCGCGT GATGAAGCAG GCCGTCGCCC ACCTGCTGCC CTACATCGAG
GCCGAGAAGC TGCTACTGGA GGCGGCCGGC GGCGATGTCA AGACCAAGGG CAAGATCATC
ATCGCCACGG TCAAGGGCGA TGTGCATGAC ATCGGCAAGA ACATCGTCAC CGTGGTCTTG
CAGTGCAACA ACTTCGAGGT CGTCAACATG GGCGTGATGG TGCCCTGCCA CGAGATTCTG
GCGCTGGCGA AAGCCGAGGG CGCGCACATC GTCGGCCTGT CGGGGCTGAT CACGCCGAGC
CTGGAAGAAA TGCAGTACGT CGCCGCCGAG ATGCAGAAGG ACGACCATTT CAGGCTCAAC
AAGATTCCGC TGATGATTGG CGGCGCCACC ACCAGCCGGG TGCACACGGC GGTGAAGATT
TCGCCGCACT ACGAAGGCCC CGTCGTCTAT GTGCCCGACG CCTCGCGCAG CGTCAGTGTG
GCGCAAAGCC TGCTGTCCGA GCAGGCGGCC AAATACATCG CCGAGCTGAA CGCCGACTAC
GACAAGGTGC GCACCCAGCA CGCCAACAAG AAGCAGACGC CGATGTGGCC GCTGGCCAAA
GCACGGGCCA ACGCCACGCC CATCGACTGG ACGAACTACA CGCCGCCCGT GCCGAAGTTC
ATCGGCCGGC GCGTGTTCAA GAATTTTGAT CTGGCCGAAC TGGCCCAGTT CATCGACTGG
GGGCCGTTCT TCCAGACCTG GGATCTGGCC GGGCCGTTCC CTGCCATCCT GACCGACGAG
GTGGTCGGCG TCGAGGCGAC GCGCGTCTAT GAAGATGCCC AGAAGATGCT CAAGCGCCTG
ATCGAAGGCC GCTGGCTCAC GGCCAGCGGC GTGATGGCGC TGCTGCCGGC CAACAGCGTC
GGCGACGACA TCGAAATCTA CACCGACGAG ACGCGCACCG AAGTGGCGAT GACCTGGCAT
GGCCTGCGCC AGCAGACCGA AAAAACGGCG GTCGATGGCG TGATGCGCCC CAGCCGCTGC
CTGGCCGATT TTGTCGCGCC CAAGGTATTA ACTCCTGAAT TGATAGCTGC TCGCACCCGT
GCAGCAAGCG CAAAAGGCCA AAATGACTTG AAAATTGCCG ATTACATCGG GGTTTTCGCG
GTCACTGCCG GCCTGGGTGC CGACAAGAAG GAAAAGGCTT TCCAGGCCGA CCATGACGAT
TATTCGTCGA TCATGTTCAA GTCGCTCGCC GACCGTCTGG CCGAAGCCTT TGCCGAAGCC
CTGCACCAGC GCGTGCGCCG TGACTTGTGG GGCTACGCGC CGGCCGAGTC GCTGGGCCAT
GACGCCTTGA TCGCCGAGCA ATACCAGGGC ATCCGCCCCG CGCCCGGCTA CCCGGCCTGC
CCGGACCACA GCGTCAAGAA GGAAATGTTC GAGCTGCTCC ACGCCGGCGA CATCGGCATG
GCGCTGACGG AAAGCCTGGC CATGACACCG GCGGCCAGCG TCAGCGGCTT TTACCTGAGC
CATCCGCAAA GCACCTATTT CAGCGTCGGC AAGATTGGCG ACGACCAACT GCAGGATCTG
GCCAGGCGGC GCGGTGCCAA GGCCGAAGAC CTGGCCAGGC TGCTGGCGCC TAATCTGTAA
 
Protein sequence
MSSIQFSHEV PPMKLSGLEP VTIGSDSLFV NIGERTNVTG SKAFARMILN GDYEQALTVA 
RQQVENGAQI IDINMDEAML DSQAAMVRFL NLIAGEPDIA RVPIMIDSSK WSVIEAGLRC
IQGKGIVNSI SMKEGVDAFK HQAKLLKRYG AAAVVMAFDE KGQADTYERK ISICERAYRV
LVDEIGFPPE DIIFDPNIFA IATGIDEHNN YAVDFIEATR WIKANLPGAK VSGGVSNVSF
SFRGNDPVRE AIHTVFLYHA IQAGMDMGIV NAGMMGVYDE LEPVLRERVE DVVLNRQPVY
KPGEAHLTPG ERLIEVAETA KSGARDDSKK YEWRALPIRA RLSHALVHGN NEFITEDTEE
VWQAIKAEGG RPLHVIEGPL MDGMNVVGDL FGQGKMFLPQ VVKSARVMKQ AVAHLLPYIE
AEKLLLEAAG GDVKTKGKII IATVKGDVHD IGKNIVTVVL QCNNFEVVNM GVMVPCHEIL
ALAKAEGAHI VGLSGLITPS LEEMQYVAAE MQKDDHFRLN KIPLMIGGAT TSRVHTAVKI
SPHYEGPVVY VPDASRSVSV AQSLLSEQAA KYIAELNADY DKVRTQHANK KQTPMWPLAK
ARANATPIDW TNYTPPVPKF IGRRVFKNFD LAELAQFIDW GPFFQTWDLA GPFPAILTDE
VVGVEATRVY EDAQKMLKRL IEGRWLTASG VMALLPANSV GDDIEIYTDE TRTEVAMTWH
GLRQQTEKTA VDGVMRPSRC LADFVAPKVL TPELIAARTR AASAKGQNDL KIADYIGVFA
VTAGLGADKK EKAFQADHDD YSSIMFKSLA DRLAEAFAEA LHQRVRRDLW GYAPAESLGH
DALIAEQYQG IRPAPGYPAC PDHSVKKEMF ELLHAGDIGM ALTESLAMTP AASVSGFYLS
HPQSTYFSVG KIGDDQLQDL ARRRGAKAED LARLLAPNL