Gene Rleg_5673 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagRleg_5673 
Symbol 
ID8016899 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameRhizobium leguminosarum bv. trifolii WSM1325 
KingdomBacteria 
Replicon accessionNC_012853 
Strand
Start bp251020 
End bp252327 
Gene Length1308 bp 
Protein Length435 aa 
Translation table11 
GC content58% 
IMG OID644827826 
Productextracellular solute-binding protein family 1 
Protein accessionYP_002979026 
Protein GI241518398 
COG category[G] Carbohydrate transport and metabolism 
COG ID[COG1653] ABC-type sugar transport system, periplasmic component 
TIGRFAM ID[TIGR01409] Tat (twin-arginine translocation) pathway signal sequence 


Plasmid Coverage information

Num covering plasmid clones15 
Plasmid unclonability p-value0.0285739 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones30 
Fosmid unclonability p-value0.0925972 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGACCAACA GCAGTGGATT CACGCATGAG ATACGTCGCC GCACCCTGCT TGCGGGTGCG 
GGAGGTGCCG CATTGATGGC GTTGGTGGGC AGCGCCAAAG CGCAGGATAT CGTCAGTGGC
GAACTTGTGG TGCTGAATTG GCTCGGCGGT TCCGAGCTCG ACATGATGCA TAAGATCCAG
ACTGCATTTA CCGCGAAATA TCCGAAAGTG ACGATCAGGG AAGTTGCCAT TACCGGCCAA
GGAGACATGC GCGGTGGCAT CCGCACGGCA CTCATGGGCG GCGAAGTGGT CGATGTCCTC
TACAACACCT GGCCGGCCTT CCGGAAGGAA CTGCTCGATG CCGGCATGCT GCGGCCGATC
GACGACCAAT GGAAGTCCTT CGGCTGGGAC AGGCTCATCA GCCAATCCTG GAAGGATCTT
GGCGCTATAG CCGGCAAGAC CTATGGCCTG ACCTACACCT TCGGCGACCG CTCCGGCATC
TGGTACAAGA AGGAGCACCT TGCCAGGGCG GGCATCACCG AGTCGCCGAG GAGCTGGGAC
GAGTTCGTCG CCAGCTTTGC GAAGCTTACA AAAGCTGGGT TCGCGGCTCC GGTCGCTATT
CCCGGCAAAT ACTGGGCGCA TGCCGAATGG TTTGAAACGC TGCTGCTGAG AACGGCAGGC
GTCGAGACGG CTTCGAAGCT AGGCGCGCAC GAAATTTCGT GGACCGATCC GGCCGTGAAG
AATGCGCTTA CAAGATATGC CGAAATGTTG ACCGCGGGCT GCTGCGGAGC GCCGAACAGC
ATGCTCGCCA ACGACTGGGA CGGGGAAGCC GACCAAATCT TCCAGGCGAA TGCCAAAAAT
TACCTGCTGA TCGGCATGTG GATGAATAAC CGCGCCAAGA ACGACTACAA ACTCACCGAA
GGTAAGGATT ACGGTCTCTT CCAGTTCCCC GCCCTCGGGA TGGGTCATGA CGACACGTCG
AGCGTCGATA CCAAGGAACT GCTCGTCACG GCAAACGGCC CCAATCCGAA GGCAGCAGAC
GCCTTCCTCG ATTTCTGGAC AAGTGCCGAG GCCGCCAACA TTCTTGCCAA GAACGGCTAT
GCGTCACCAA GCAGCAATAC CGACACGTCG CTCTATGGCG AGACGCAGAA GGTGGCGACA
TCAGCGGTCG CAAGCTCGAA GCTGCAATTC GTGCTCGGAG ATCTCTTGCC CGGCGATCTC
GTCGATGAAT ATCGGGTGCA ACTGCAGAAA TTTCTCCAGG ATCCCTCGGC TGCCAATATC
GATACCGTCC TTGCGGCAAT CGAAACCAAG GCCCAGGGAT CCTACTGA
 
Protein sequence
MTNSSGFTHE IRRRTLLAGA GGAALMALVG SAKAQDIVSG ELVVLNWLGG SELDMMHKIQ 
TAFTAKYPKV TIREVAITGQ GDMRGGIRTA LMGGEVVDVL YNTWPAFRKE LLDAGMLRPI
DDQWKSFGWD RLISQSWKDL GAIAGKTYGL TYTFGDRSGI WYKKEHLARA GITESPRSWD
EFVASFAKLT KAGFAAPVAI PGKYWAHAEW FETLLLRTAG VETASKLGAH EISWTDPAVK
NALTRYAEML TAGCCGAPNS MLANDWDGEA DQIFQANAKN YLLIGMWMNN RAKNDYKLTE
GKDYGLFQFP ALGMGHDDTS SVDTKELLVT ANGPNPKAAD AFLDFWTSAE AANILAKNGY
ASPSSNTDTS LYGETQKVAT SAVASSKLQF VLGDLLPGDL VDEYRVQLQK FLQDPSAANI
DTVLAAIETK AQGSY