Gene Mlg_1856 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagMlg_1856 
Symbol 
ID4268074 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameAlkalilimnicola ehrlichii MLHE-1 
KingdomBacteria 
Replicon accessionNC_008340 
Strand
Start bp2116112 
End bp2117476 
Gene Length1365 bp 
Protein Length454 aa 
Translation table11 
GC content69% 
IMG OID638126612 
Productpeptidase RseP 
Protein accessionYP_742690 
Protein GI114321007 
COG category[M] Cell wall/membrane/envelope biogenesis 
COG ID[COG0750] Predicted membrane-associated Zn-dependent proteases 1 
TIGRFAM ID[TIGR00054] RIP metalloprotease RseP 


Plasmid Coverage information

Num covering plasmid clones14 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones52 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGGGCATCC TCTGGAGCAT ACTGGCGTTC GTCGTGGCCA TTGGCATCCT GGTCACGGTG 
CACGAGTTCG GCCATTTCTG GGTGGCCCGG CGGGCCGGCA TCAAGGTGCT GCGCTTCTCG
GTGGGCTTCG GCCGGCCGCT GTTGCGCTGG CGCCGTGGGG CGGATCGCAC CGAATACGTC
ATTGCCGCGA TCCCGCTCGG CGGCTACGTG AAGATGCTGG ACGAACGCGA GGCCGAGGTG
CCCGAGGCCG AACGCCACCG CGCCTTCAAC GTCCAGCCGC TGTACAAGCG CACCGCCGTG
GTGCTCGCCG GACCCTTGTT CAATTTCCTG TTCGCAGTGC TCGCTTACAT GGCCATCGGC
CTGCTCGGCA CTGTCGAGAT GCGCCCGGTC CTGGGGCCGG TGGCGGAGAA CACGCCGGCG
GCTGAGGCCG GTTTTCAGGA GGGCGATGAG CTGCTCGCCA TCGGCGGCCG CGAGACCCCC
ACCTGGCAGC GCACCGCCAT GGCGCTGGTG GATGCCGGCT TTCACCGCGC CGACATCCCG
GTGGAGGTCC GGGGCGAGGA CGGCCGTGCC CGCAGTCTGG TGCTGGACAT GACCCTGGCC
GGTGAGATCG GGCGGGCGGA CAATCTGCTG GCGCAGGCCG GTTTCCGTCC CTGGACCCCG
GCCTTGGACC CGGTGCTCGG CCGTGTGGTG GATGACGGGC CCGCGGCCCG GGCCGGGCTC
ATGGCCGGCG ACCGCATCGT CTCGGTGGAG GGCGAGCCGG TGGCGGAATG GCGTGAGCTG
GTCGAGTGGA TCGAGCACCA TCCGGGCGAG GTTCTGACCC TCACGATCGA GCGCGACGGC
CGTCAGGAGA CCATCGATAC GCGGCTGGAC AGCGTGGAGG CGGCCGGGCG CACCATCGGT
CAGCTTGGGG TGGCCCCCGA GGTGCCGGAG GGGGCCTATG ACCGGCTCTA CCGCGAGGTC
CAATACGGAC CGGTCGGGGC GCTGGGCCAT GGCCTGTCCT CCACCTGGGA TGCCAGCGTG
CTGACGGTGA AGATCCTCGG CCGTATGGTG ATCGGCCAGG CCTCGCTGCA GAATCTTAGC
GGCCCGCTCA CCATCGGGCA GTTTGCGGGC GATACCGCCT CGCTGGGCGT GGTACCCTTC
CTGGGCTTCC TCGCCATCGT CAGTATCAGT CTGGGGATCA TCAACCTGTT GCCGATCCCC
ATCCTGGACG GCGGGCACTT GCTCTATTTC GCGGTCGAGG CCGTACGCGG CAAGCCGCTG
TCGGAGTACG CCCAGGCGGT GGGCCAGCAG GTGGGGCTGC TGATGCTGTT CCTGCTCATG
GGACTGGCGT TCTACAACGA CCTGGCGCGC CTGTTCGGCG GCTAA
 
Protein sequence
MGILWSILAF VVAIGILVTV HEFGHFWVAR RAGIKVLRFS VGFGRPLLRW RRGADRTEYV 
IAAIPLGGYV KMLDEREAEV PEAERHRAFN VQPLYKRTAV VLAGPLFNFL FAVLAYMAIG
LLGTVEMRPV LGPVAENTPA AEAGFQEGDE LLAIGGRETP TWQRTAMALV DAGFHRADIP
VEVRGEDGRA RSLVLDMTLA GEIGRADNLL AQAGFRPWTP ALDPVLGRVV DDGPAARAGL
MAGDRIVSVE GEPVAEWREL VEWIEHHPGE VLTLTIERDG RQETIDTRLD SVEAAGRTIG
QLGVAPEVPE GAYDRLYREV QYGPVGALGH GLSSTWDASV LTVKILGRMV IGQASLQNLS
GPLTIGQFAG DTASLGVVPF LGFLAIVSIS LGIINLLPIP ILDGGHLLYF AVEAVRGKPL
SEYAQAVGQQ VGLLMLFLLM GLAFYNDLAR LFGG