Gene TM1040_0431 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagTM1040_0431 
Symbol 
ID4076191 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameRuegeria sp. TM1040 
KingdomBacteria 
Replicon accessionNC_008044 
Strand
Start bp442988 
End bp444295 
Gene Length1308 bp 
Protein Length435 aa 
Translation table11 
GC content61% 
IMG OID638005726 
Productextracellular solute-binding protein 
Protein accessionYP_612426 
Protein GI99080272 
COG category[G] Carbohydrate transport and metabolism 
COG ID[COG1653] ABC-type sugar transport system, periplasmic component 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones24 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones17 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGTACCTGA GAAACGCACT TTGTGCGGCC TCTGCACTTG CGGTTATGGC AACTGGCGCC 
GTCCAGGCCG AGACCACATT GACCATCGCG ACCGTGAACA ACGGCGACAT GATCCGGATG
CAGGGCCTCA CTGACGACTT CACCGCCAAA CATCCCGACA TCCAGCTGGA GTGGGTCACG
CTCGAAGAGA ACGTGCTGCG TCAGCGCGTG ACACAAGACA TCGCCACCAA CGGCGGCCAG
TTCGATGTGA TGACCATCGG CATGTACGAA ACCCCGATCT GGGCCGCTCA GAATTGGCTT
GTGCCGCTGA CCGACATGGG GGCTGATTAC GACGCGGACG ACATCCTGCC CGCGATGCGC
GCTGGCCTTT CACACAATGG CACTCTCTAT GCCGCGCCGT TTTACGGTGA AAGCTCCATG
GTCATGTATC GCACCGACCT GATGGAGGCC GCGGGCCTGA CAATGCCCGA AGCCCCCACA
TGGGAGTTCA TCAAGGAAGC CGCCGCCGCG ATGACAGACA AGGACGCCGA AATCTACGGC
GCATGCCTGC GCGGCAAGGC AGGATGGGGC GAGAACATGG CCTTCATCAC CACCGTGGCG
AACAGCTTTG GTGCGCGCTG GTTTGACGAA GACTGGACGC CGCAGCTCGA CAGCCCCGAG
TGGAAAGAGG CGGTCACTTT CTACAACGAT CTGCTGCAAA GCTACGGACC TCCGGGTGCC
TCTACAAATG GGTTCAACGA GAACCTCGCA CTGTTCCAGC AAGGCAAGTG TGGCATGTGG
ATCGACGCGA CTGTGGCGGC CTCCTTCGTG ACCAACCCCG ACGATTCCAC CGTCGCTGAC
AAGGTTGGAT TTGCTCTGGC ACCCGACACC GGCAAGGGCA AACGGGCCAA CTGGCTCTGG
GCCTGGGCGC TGGCCGTGCC TGCGGGTTCT GACGCCAAGG ATGAGGCCAA GGCATTCATC
GAATGGGCCA CCTCCAAGGA GTATCTTGCG CTTGTGGCGG AAAACGAAGG TTGGGCCAAT
GTACCGCCTG GCAGCCGCAC GTCGCTTTAT GAGAACCCTG AATACGCCAA GGTTCCCTTT
GCGCAGATGA CGCTCGACTC GATCAATGCG GCGGACCCCA ACAGCCCCAC CGTGGATCCC
GTGCCCTATG TGGGCATCCA GTATGTCGCG ATCCCCGAAT GGGCCGGCAT CGGCACCAGC
GCAGGCCAGG AATTCTCGGC CATGGTCGCA GGTCAGCAAA CCCCGGACGA AGCGCTTGCA
AAAGCACAGG CTCTGGTCGC TGACGAAATG GAAGCCGCAG GCTACTAA
 
Protein sequence
MYLRNALCAA SALAVMATGA VQAETTLTIA TVNNGDMIRM QGLTDDFTAK HPDIQLEWVT 
LEENVLRQRV TQDIATNGGQ FDVMTIGMYE TPIWAAQNWL VPLTDMGADY DADDILPAMR
AGLSHNGTLY AAPFYGESSM VMYRTDLMEA AGLTMPEAPT WEFIKEAAAA MTDKDAEIYG
ACLRGKAGWG ENMAFITTVA NSFGARWFDE DWTPQLDSPE WKEAVTFYND LLQSYGPPGA
STNGFNENLA LFQQGKCGMW IDATVAASFV TNPDDSTVAD KVGFALAPDT GKGKRANWLW
AWALAVPAGS DAKDEAKAFI EWATSKEYLA LVAENEGWAN VPPGSRTSLY ENPEYAKVPF
AQMTLDSINA ADPNSPTVDP VPYVGIQYVA IPEWAGIGTS AGQEFSAMVA GQQTPDEALA
KAQALVADEM EAAGY