Gene Dole_1126 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagDole_1126 
Symbol 
ID5693960 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameDesulfococcus oleovorans Hxd3 
KingdomBacteria 
Replicon accessionNC_009943 
Strand
Start bp1335747 
End bp1338392 
Gene Length2646 bp 
Protein Length881 aa 
Translation table11 
GC content62% 
IMG OID641263720 
ProductDNA mismatch repair protein MutS 
Protein accessionYP_001529010 
Protein GI158521140 
COG category[L] Replication, recombination and repair 
COG ID[COG0249] Mismatch repair ATPase (MutS family) 
TIGRFAM ID[TIGR01070] DNA mismatch repair protein MutS 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.0124737 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGGCTTCCA CAGGCGCCAC TCCCATGATG CAGCAGTATC TCTCCATCAA GGAGCAGCAC 
CGGGACGCCA TTCTTTTTTA CCGAATGGGC GACTTTTACG AGATGTTTTT TGAGGACGCT
CAAACCGCGG CCCCGGTCCT TGAGATCGCT CTGACCTCCC GCAACAAGAA CGACACCGAT
CCCATTCCCA TGTGCGGTGT GCCGGTAAAG GCCGCGGACG GCTATATCGG CCGGCTCATC
GAAAACGGGT TCAAGGTGGC GGTATGCGAG CAGACCGAGG ACCCTGCCGC GGCCAAAGGC
CTGGTCCGGC GGGACGTGGT GCGCATCGTC ACTCCGGGCA TGATCATCGA CAATGCTCTG
CTGGAAAAGG GAACCAATAA CTACGTTGTC TGCCTGGCCC ATGCCGACGG TGTTGTGGGG
TTTGCCAGCG TGGATATCTC CACCGGCACT TTTCGGGTGT GCGAGTCCTC CGACCTGCGG
GCCGTGCGCC ACGAGCTGCT GCGCATCGCG CCCCGGGAAG TGGTAATACC GGAATCCGGC
GCCGATGACG CGGCGCTTTC GCCCTTTGTT TCCCTTTTTC CGCCGGCCAT TCGAACAACG
CTCGCTAACC GGGAGTTTGA TTACAGAACC GCCTGCCAGC GGCTGACCGA CCAGTTTCAG
ACCCGGTCCC TGGAGGGGTT CGGGTGCCGG GGCCTCAAAC CCGGCATTGT CGCGGCCGGG
GCCCTGCTTT CCTATGTAAA CGATACCCAG AGACAGAAGG CGTCCCACCT GACCGGGCTG
GAGGTCTACA GCATCGACCA GTACCTGCTG ATGGACGAGG TGACCTGCCG GAATCTGGAA
CTGGTGGCCA ACCTTCGCAA CAATGGCAGG CAGGGAACCC TTATTGATGT GCTGGACGCC
TGCGTCACCG CCATGGGCAG CCGCCTGCTG CGGCGCTGGA TGCTCTATCC CCTGCTGTCG
GCAGAAGCCA TCAACCGGCG GCTGGACGCG GTGGCAGAGG CCAAAGAGGG CCTGGGCACT
CGAAAGGCGG TGCGGGAACT GCTCAAACAG GTCTACGATA TCGAGCGGCT TACCAGCCGG
GCCGTTATGG GCCGGGTCAC CCCTCGGGAC CTGCTGGCCT TGAAACAGAC CCTTTTCGCC
CTGCCGGGTC TGGCAACAGA ACTGAAGTCT TTTGACAGCC CTTTTTTCTC CTTTGCCGGG
GAACCGGGGC CCGAAGGCCT TGATAAGCTG GCCGGCCTGG CCGATCTGCT GAAGGCGGCG
GTGCGGGAGG ACGCGCCGGT TTCCATCGCT GACGGCGGTG TCATCAACCC CGACTATCAT
CCCCGGCTGG CCGAACTGGT AACCATCAGC CGGGACGGCA AGAGCAGCCT GGCCCGGCTG
GAGGCAACGG AAAAAGAGAA GACCGGCATT TCCACCCTCA AGGTGCGGTA CAACAAGGTG
TTTGGTTACT ATATCGAGGT ACCCCGGTCC CAGGTGGGGG CCGTGCCGGC TCACTACGTT
CGCAAGCAGA CCCTTGTCAA CGGTGAGCGC TACATCACCG ACGAGCTTAA GGTGTTTGAG
GAAAAAGCCC TGGGCGCCGA AGAACAGCGC GTTCGGCTGG AGCAGGAGTT GTTTGCCGAT
ATCGTGGGCC GGGTGACCGC GTGCAGCCCG ATGCTGTTTG CCGTGGCCCG GGTCGCGGCC
GGAATCGACG TGTTGTGCGC CCTGGCCCAG GTGGCCGATG ACCATGACTA TGTCCGGCCC
GAGATGCTGT CCGGCGGCGA GATCATCATT GAAGAGGGCC GTCATCCCGT GGTGGAGCGC
ATGCTTTCCG GCGAACGGTA CGTGCCCAAC AGCATTACGT TAAACGATAC CGACCGGCAG
CTGCTGATCA TCACCGGTCC CAACATGGCG GGCAAATCCA CGGTGCTGCG CAAGGTGGCG
CTGTTTTCGG TCATGGCCCA GATGGGCTCC TTTGTACCGG CCCGGCGGGC CGCCATGGGT
GTGGTGGACC GGCTCTTTAC CCGGGTGGGG GCCCTGGACA ACCTGGCCTC AGGCCAGAGT
ACCTTCATGG TGGAGATGGA AGAGACGGCC AACATCATCA ACAACGCCAC GCCGAAAAGC
CTGGTGGTGA TCGACGAGAT CGGCCGGGGC ACCAGCACCT ACGACGGCCT GAGCATTGCC
TGGGCCGTGG CCGAGGCCCT GCATGATCTG CACGGCAGAG GGGTCAAGAC CCTGTTTGCC
ACCCATTACC ACGAGCTGAC CGAACTGGAA AACACCCGGC CCCGGGTGAA GAACTTTCAT
ATTGCCGTCA AGGAGTGGAA CGATACCATC ATTTTTTTAA GAAAGCTGGT GGAGGGCAGC
ACCAACCGCA GCTACGGCAT TCAGGTGGCA AGGCTGGCCG GCATTCCCGG CCCGGTGATC
GCCAGGGCCA AGAAGATTCT GCTGGACATC GAGCAGGGCA CCTACAGTTT TGAGGCAAAG
TCCGGCACTG CTCCGGGCAC CGGACAGAGC GGCCCGGTTC AGCTCTCCCT GTTTACCCCG
CCGGAACAGA TGCTGGTGGA CCGGCTTCAA AAGGTCGACA TTTCAACCAT GACGCCCCTG
GAGGCATTGA ACTGCCTTCA CGAACTGCAA CAGAAGGCGC ACGCCATATC GGAGACCGAC
GGATGA
 
Protein sequence
MASTGATPMM QQYLSIKEQH RDAILFYRMG DFYEMFFEDA QTAAPVLEIA LTSRNKNDTD 
PIPMCGVPVK AADGYIGRLI ENGFKVAVCE QTEDPAAAKG LVRRDVVRIV TPGMIIDNAL
LEKGTNNYVV CLAHADGVVG FASVDISTGT FRVCESSDLR AVRHELLRIA PREVVIPESG
ADDAALSPFV SLFPPAIRTT LANREFDYRT ACQRLTDQFQ TRSLEGFGCR GLKPGIVAAG
ALLSYVNDTQ RQKASHLTGL EVYSIDQYLL MDEVTCRNLE LVANLRNNGR QGTLIDVLDA
CVTAMGSRLL RRWMLYPLLS AEAINRRLDA VAEAKEGLGT RKAVRELLKQ VYDIERLTSR
AVMGRVTPRD LLALKQTLFA LPGLATELKS FDSPFFSFAG EPGPEGLDKL AGLADLLKAA
VREDAPVSIA DGGVINPDYH PRLAELVTIS RDGKSSLARL EATEKEKTGI STLKVRYNKV
FGYYIEVPRS QVGAVPAHYV RKQTLVNGER YITDELKVFE EKALGAEEQR VRLEQELFAD
IVGRVTACSP MLFAVARVAA GIDVLCALAQ VADDHDYVRP EMLSGGEIII EEGRHPVVER
MLSGERYVPN SITLNDTDRQ LLIITGPNMA GKSTVLRKVA LFSVMAQMGS FVPARRAAMG
VVDRLFTRVG ALDNLASGQS TFMVEMEETA NIINNATPKS LVVIDEIGRG TSTYDGLSIA
WAVAEALHDL HGRGVKTLFA THYHELTELE NTRPRVKNFH IAVKEWNDTI IFLRKLVEGS
TNRSYGIQVA RLAGIPGPVI ARAKKILLDI EQGTYSFEAK SGTAPGTGQS GPVQLSLFTP
PEQMLVDRLQ KVDISTMTPL EALNCLHELQ QKAHAISETD G