Gene Smed_3943 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagSmed_3943 
Symbol 
ID5318524 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameSinorhizobium medicae WSM419 
KingdomBacteria 
Replicon accessionNC_009620 
Strand
Start bp392457 
End bp394262 
Gene Length1806 bp 
Protein Length601 aa 
Translation table11 
GC content60% 
IMG OID640775753 
ProductPyrrolo-quinoline quinone 
Protein accessionYP_001312686 
Protein GI150376090 
COG category[G] Carbohydrate transport and metabolism 
COG ID[COG4993] Glucose dehydrogenase 
TIGRFAM ID[TIGR03075] PQQ-dependent dehydrogenase, methanol/ethanol family 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.237421 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones26 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGAAACGAC TTCTTACTAT GCTTGCCTTC ATGTCGATAG GCGGCGGTGC GCAGGTCGCA 
TTCGCCAACG ATGATTTGCA GAAGCTCATC GATGACCCCA ACCAATGGGC GATCCAGACC
GGCGACTATG CCAACCTGCG CTATTCCAAG CTTGACCAGA TCAACAAGGA CAACGTAGGC
AAACTCCAGG TCGCCTGGAC CTTCTCGACC GGCGTTCTGC GCGGTCACGA AGGTTCGCCG
CTTGTGATCG GCGACATGAT GTATGTGCAT ACGCCGTTCC CGAACACCGT CTATGCGCTG
GATCTCAGCA AAGACGGCCA GATCGTCTGG AAATACGAGC CGAAGCAGGA TCCGAACGTG
ATCCCGGTGA TGTGTTGCGA TACGGTCAAT CGCGGCGTCG CCTATGCGGA TAACAAGATC
TTTCTGCATC AGGCCGATAC GACGGTCGTG GCACTCGATG CCAAGACCGG CAAGGTGGCA
TGGAGCGTCA AGAACGGCGA TGCGACCAAG GGCGAGACCA ACACCGCCAC CGTCATGCCT
GTCAAGGACA AGGTTCTCGT CGGCATATCC GGCGGCGAGT TCGGCGTTCG CGGCCATGTG
ACCGCTTATT CGATCGCGGA TGGCAAGCTC CTGTGGCGCG CTTACTCCAT GGGGCCGGAC
AGCGACACGC TGATGAATCC TGAAAAGACG ACCCACCTTG GCAAGCCGGT CGGAAAGGAC
TCAGGTCTGA CGACCTGGGA GGGCGATCAG TGGAAGATCG GGGGCGGTAC GACCTGGGGC
TGGTACTCCT ATGATCCCGA ACTCAACATG GTGTATTACG GAACCGGCAA CCCATCCACC
TGGAACCCGA CCCAGCGGCC TGGCGACAAT CGCTGGTCGA TGACGGTCTT CGCGCGGGAC
GTCGATACCG GGGAGGCCAA ATGGGTCTAC CAGATGACGC CCCACGACGA ATGGGACTAT
GACGGCGTCA ACGAGATGAT CCTGACCGAA CAGCAGATCG ACGGCAAGGA TCGCAAACTG
CTGACCCACT TCGACCGCAA CGGCTTCGGC TATACGATGG ATCGCGTCAC GGGCGAGTTG
CTCGTAGCGG AGAAATACGA TCCCACCGTC AACTGGGCGA CTGAAGTGGT CATGGACTCC
AAGTCGGAAC AGTTTGGCCG GCCGCAAGTC GTAGCTCAGT ATTCGACTGA GCAGAACGGC
GAAGACACCA ACACGACCGG CGTCTGCCCG GCCGCTCTCG GTACCAAGGA CCAGCAGCCG
GCGGCCTATT CGCCGAAAAC CGAGCTGTTC TACGTGCCTA CCAACCACGT CTGCATGGAC
TACGAACCCT TCCGGGTAAG CTACACCGCC GGCCAGCCCT ATGTAGGGGC CACGCTGTCG
ATGTACCCGC CGAAGGATAG CCATGGCGGC ATGGGCAACT TCATCGCCTG GGACAACAAG
GAAGGCAAAA TCAAGTGGTC CCTACCGGAG CCGTTCTCGG TCTGGTCCGG CGCTCTGGCG
ACAGCCGGCG ACGTGGTCTT CTACGGAACG CTCGAAGGTT ATCTGAAGGC GGTCGACGCC
GCGACCGGCA AGGAGCTCTA TCGCTTCAAG ACGCCTTCCG GAGTGATTGG CAACGTGATG
ACCTATGCCC GCGAAGGCAA GCAGTACGTG GCCGTCCTCT CGGGTGTTGG GGGCTGGGCC
GGAATCGGTC TGGCAGCGGG CCTGACCAAC CCAACTGAAG GTCTCGGAGC CGTTGGCGGC
TACTCGGCCC TCAGCAACTA CACCGCGCTC GGCGGTACGC TCACCGTGTT CAAGCTGCCG
GAATAA
 
Protein sequence
MKRLLTMLAF MSIGGGAQVA FANDDLQKLI DDPNQWAIQT GDYANLRYSK LDQINKDNVG 
KLQVAWTFST GVLRGHEGSP LVIGDMMYVH TPFPNTVYAL DLSKDGQIVW KYEPKQDPNV
IPVMCCDTVN RGVAYADNKI FLHQADTTVV ALDAKTGKVA WSVKNGDATK GETNTATVMP
VKDKVLVGIS GGEFGVRGHV TAYSIADGKL LWRAYSMGPD SDTLMNPEKT THLGKPVGKD
SGLTTWEGDQ WKIGGGTTWG WYSYDPELNM VYYGTGNPST WNPTQRPGDN RWSMTVFARD
VDTGEAKWVY QMTPHDEWDY DGVNEMILTE QQIDGKDRKL LTHFDRNGFG YTMDRVTGEL
LVAEKYDPTV NWATEVVMDS KSEQFGRPQV VAQYSTEQNG EDTNTTGVCP AALGTKDQQP
AAYSPKTELF YVPTNHVCMD YEPFRVSYTA GQPYVGATLS MYPPKDSHGG MGNFIAWDNK
EGKIKWSLPE PFSVWSGALA TAGDVVFYGT LEGYLKAVDA ATGKELYRFK TPSGVIGNVM
TYAREGKQYV AVLSGVGGWA GIGLAAGLTN PTEGLGAVGG YSALSNYTAL GGTLTVFKLP
E