Gene Rleg_1638 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagRleg_1638 
Symbol 
ID8012709 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameRhizobium leguminosarum bv. trifolii WSM1325 
KingdomBacteria 
Replicon accessionNC_012850 
Strand
Start bp1629792 
End bp1630790 
Gene Length999 bp 
Protein Length332 aa 
Translation table11 
GC content62% 
IMG OID644824224 
ProductTRAP dicarboxylate transporter- DctP subunit 
Protein accessionYP_002975465 
Protein GI241204369 
COG category[G] Carbohydrate transport and metabolism 
COG ID[COG1638] TRAP-type C4-dicarboxylate transport system, periplasmic component 
TIGRFAM ID[TIGR00787] tripartite ATP-independent periplasmic transporter solute receptor, DctP family
[TIGR01409] Tat (twin-arginine translocation) pathway signal sequence 


Plasmid Coverage information

Num covering plasmid clones20 
Plasmid unclonability p-value0.503152 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones18 
Fosmid unclonability p-value0.0466092 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGGACAATT TCAACCGACG CAATTTCCTG AAGACCGCAG CTCTCGCCGG AACGGCGCTT 
GCAGCGCCCG CCTTCGTCCG CACGGCCGCC GCCCGCACGA CGACGATCAC GATCGCCTCG
CTGCTCGGTG ACGACAAGCC GGAAACGAAG ATCTGGGTGA AGATCGGCGA ACTGGTCGAA
GCGAAACTCC CCGGCCAGTT CAAGTTCAAT ATCGTCAGGA ACGGCGCGCT CGGCGGCGAG
AAGGAAGTGG CCGAAGGCGT GCGCCTCGGC TCCATCCAGG CCAGCCTCTC GACGGTCTCG
TCGCTCTCCG GCTGGGCGCC GGAACTGCAG ATCCTCGATC TGCCCTTTCT CTTCCGCGAT
GCCGACCATG TTCGCAGGAC CGTCGCCGGT GATGTCGGCG CCGATCTCAA GCAGAAACTG
CAGGCGCAGA ATTTCGTCGT CGGCGATTTC ATCAATTACG GCGCCCGCCA TCTCCTGACC
AAGGAGCCGG TGACGCGGCC AGAACAACTC AAGGGCAAGC GCATTCGCGT CATCCAGAGC
CCTCTGCACA CCAAGCTCTG GAGCGCATTC GGCACGACGC CGATCGGCAT TCCGATCACC
GAGACCTATA ATGCGCTCGC AACCGGCGTC GCCGACGCGA TGGATCTGAC CAAGTCGGCC
TATGCCGGCT TCAAGCTCTA CGAGGTCGTG CCCGATATGA CCGAGACCGG CCACATCTGG
GCATCCGGCG TCATCTATTA TGCCTCGACC TTCTGGGCCG GCCTCAATGA CGAGCAGAAG
GCGGTGTTCC AGCAGGCCTC CAGCGAGGGG GCTGCCTATT TCAACCAGTT GATCGTCGAT
GACGAATCCA AATCCGTCGA GACGGCGCTT GCCAACGGCG GAAAACTCTT GAAGCCGGAA
GCCTTCGACG AATGGCAGAA GGGCGCCCAG GGGGTGTGGG ACGATTTCGC GCCGGTTGTC
GGCGGCATCG ACAGGATCAA ATCGATCCAG GCCGCCTGA
 
Protein sequence
MDNFNRRNFL KTAALAGTAL AAPAFVRTAA ARTTTITIAS LLGDDKPETK IWVKIGELVE 
AKLPGQFKFN IVRNGALGGE KEVAEGVRLG SIQASLSTVS SLSGWAPELQ ILDLPFLFRD
ADHVRRTVAG DVGADLKQKL QAQNFVVGDF INYGARHLLT KEPVTRPEQL KGKRIRVIQS
PLHTKLWSAF GTTPIGIPIT ETYNALATGV ADAMDLTKSA YAGFKLYEVV PDMTETGHIW
ASGVIYYAST FWAGLNDEQK AVFQQASSEG AAYFNQLIVD DESKSVETAL ANGGKLLKPE
AFDEWQKGAQ GVWDDFAPVV GGIDRIKSIQ AA