Gene Sbal195_1199 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagSbal195_1199 
Symbol 
ID5752926 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameShewanella baltica OS195 
KingdomBacteria 
Replicon accessionNC_009997 
Strand
Start bp1422454 
End bp1423554 
Gene Length1101 bp 
Protein Length366 aa 
Translation table11 
GC content44% 
IMG OID641287468 
Productextracellular solute-binding protein 
Protein accessionYP_001553634 
Protein GI160874318 
COG category[E] Amino acid transport and metabolism 
COG ID[COG0687] Spermidine/putrescine-binding periplasmic protein 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.000384833 
Plasmid hitchhikingNo 
Plasmid clonabilitydecreased coverage 
 

Fosmid Coverage information

Num covering fosmid clones17 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGAAGCTAT TTAATAAGAT GACCACTCTA GCTCTGGTAA CTGCGAGCGT ATTAGCGAGC 
GCAGCGGCCC AAGCGGAAGA AGTGGTTCGC GTGTATAACT GGTCAGATTA TATCGCGGAA
GATACCTTAG AAAACTTCAA GAAAGAAACG GGCATTCGGG TTATTTACGA TGTGTTCGAT
AGTAACGAAG TGCTTGAGGC TAAATTATTG TCTGGTCGAA GTGGCTACGA TATTGTTGTC
CCTTCTAACC ACTTTCTCGC TAAGCAAATC AAAGCGGGTG CTTTCAAACC TTTAGACCGC
GCTAAGCTAT CTAATTTCAA AAATTTAAAT CCCGCCCTGA TGAAGCTACT TGAGAAAGCC
GATCCGGGTA ACCAGTATGC AGTGCCTTAT TTATGGGGAA CCAATGGTAT TGGTTACAAC
ATCGATAAAG TGAAAGCGGC TGTGGGTGAA GATGCGCCAT TCAACTCAAT GGAACTGATC
TTCAATCCTA AATATGCTGA AAAAATCTCT AAGTGTGGCT TTGCTATGCT GGACTCTGCC
GACGATATGG TGCCCCAAGC ACTGATTTAT TTAGGTTTAG ATCCTAACAG TTCCAACCCA
AGCGATTATG AAAAAGCCGG TGAGTTACTG GAAAAAATCC GTCCTTACGT GACCTATTTC
CACTCATCTC GCTATATTTC TGACTTAGCA AACGGTGACA TTTGTGTGGC CTTTGGTTTT
TCTGGTGACG TATTCCAAGC TAAAGCGCGT GCTGAAGAGG CGGGTAATGG CAATAAGATT
GGTTACTCGA TTCCAAAAGA AGGCGCTAAC CTGTGGTTTG ATATGTTAGC TATCCCAGCC
GATTCGACTA ACGCAGATAA TGCACTGACG CTGATTAACT ATTTCCTCCG TCCAGAAGTC
ATAGCGCCTA TCTCTAACTA TGTGGCCTAT GCTAACCCGA ACGATCCTGC ACAACCTCTG
GTTGATGAGG CTATCCGCAC CGATCCCGCG ATTTATCCAC CGCAAGAAGT GTTAGATAAA
CTTTATATTG GTGAAATCCG TCCTTTGAAA ATCCAACGCG TATTAACCCG TGTTTGGACC
AAAGTGAAGT CAGGACAATA G
 
Protein sequence
MKLFNKMTTL ALVTASVLAS AAAQAEEVVR VYNWSDYIAE DTLENFKKET GIRVIYDVFD 
SNEVLEAKLL SGRSGYDIVV PSNHFLAKQI KAGAFKPLDR AKLSNFKNLN PALMKLLEKA
DPGNQYAVPY LWGTNGIGYN IDKVKAAVGE DAPFNSMELI FNPKYAEKIS KCGFAMLDSA
DDMVPQALIY LGLDPNSSNP SDYEKAGELL EKIRPYVTYF HSSRYISDLA NGDICVAFGF
SGDVFQAKAR AEEAGNGNKI GYSIPKEGAN LWFDMLAIPA DSTNADNALT LINYFLRPEV
IAPISNYVAY ANPNDPAQPL VDEAIRTDPA IYPPQEVLDK LYIGEIRPLK IQRVLTRVWT
KVKSGQ