Gene EcE24377A_4586 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagEcE24377A_4586 
SymbolmalE 
ID5587010 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameEscherichia coli E24377A 
KingdomBacteria 
Replicon accessionNC_009801 
Strand
Start bp4584347 
End bp4585537 
Gene Length1191 bp 
Protein Length396 aa 
Translation table11 
GC content51% 
IMG OID640928203 
Productmaltose ABC transporter periplasmic protein 
Protein accessionYP_001465535 
Protein GI157156613 
COG category[G] Carbohydrate transport and metabolism 
COG ID[COG2182] Maltose-binding periplasmic proteins/domains 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones37 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGAAAATAA AAACAGGTGC ACGCATCCTC GCATTATCCG CATTAACGAC GATGATGTTT 
TCCGCCTCGG CTCTCGCCAA AATCGAAGAA GGTAAACTGG TAATCTGGAT TAACGGCGAT
AAAGGCTATA ACGGTCTCGC TGAAGTCGGT AAGAAATTCG AGAAAGATAC CGGAATTAAA
GTCACCGTTG AGCATCCGGA TAAACTGGAA GAGAAATTCC CACAGGTTGC GGCAACTGGC
GATGGCCCTG ACATTATCTT CTGGGCGCAC GACCGCTTTG GTGGCTACGC TCAATCTGGC
CTGTTGGCTG AAATCACCCC GGACAAAGCG TTCCAGGACA AGCTGTATCC GTTTACCTGG
GATGCCGTAC GTTACAACGG CAAGCTGATT GCTTACCCGA TCGCTGTTGA AGCGTTATCG
CTGATTTATA ACAAAGATCT GCTGCCGAAC CCGCCAAAAA CCTGGGAAGA GATCCCGGCG
CTGGATAAAG AACTGAAAGC GAAAGGTAAG AGCGCGCTGA TGTTCAACCT GCAAGAACCG
TACTTCACCT GGCCGCTGAT TGCTGCTGAC GGGGGTTATG CGTTCAAGTA TGAAAATGGC
AAATACGACA TTAAAGACGT GGGCGTGGAT AACGCTGGCG CGAAAGCGGG TCTGACCTTC
CTGGTTGACC TGATTAAAAA CAAACACATG AATGCAGACA CCGATTACTC CATCGCAGAA
GCTGCCTTTA ATAAAGGCGA AACAGCGATG ACCATCAACG GCCCGTGGGC ATGGTCCAAC
ATCGACACCA GCAAAGTGAA TTATGGTGTA ACGGTACTGC CGACCTTCAA GGGTCAACCA
TCTAAACCGT TCGTTGGCGT GCTGAGCGCC GGTATTAACG CCGCCAGTCC GAACAAAGAG
CTGGCGAAAG AGTTCCTCGA AAACTATCTG CTGACTGACG AAGGTCTGGA AGCGGTTAAT
AAAGACAAAC CGCTGGGTGC CGTAGCGCTG AAGTCTTACG AGGAAGAGTT GGCGAAAGAT
CCACGTATTG CAGCCACCAT GGAAAACGCC CAGAAAGGTG AAATCATGCC GAACATCCCG
CAGATGTCCG CTTTCTGGTA TGCCGTGCGT ACTGCGGTGA TCAACGCCGC CAGCGGTCGT
CAGACTGTCG ATGAAGCCCT GAAAGACGCG CAGACTCGTA TCACCAAGTA A
 
Protein sequence
MKIKTGARIL ALSALTTMMF SASALAKIEE GKLVIWINGD KGYNGLAEVG KKFEKDTGIK 
VTVEHPDKLE EKFPQVAATG DGPDIIFWAH DRFGGYAQSG LLAEITPDKA FQDKLYPFTW
DAVRYNGKLI AYPIAVEALS LIYNKDLLPN PPKTWEEIPA LDKELKAKGK SALMFNLQEP
YFTWPLIAAD GGYAFKYENG KYDIKDVGVD NAGAKAGLTF LVDLIKNKHM NADTDYSIAE
AAFNKGETAM TINGPWAWSN IDTSKVNYGV TVLPTFKGQP SKPFVGVLSA GINAASPNKE
LAKEFLENYL LTDEGLEAVN KDKPLGAVAL KSYEEELAKD PRIAATMENA QKGEIMPNIP
QMSAFWYAVR TAVINAASGR QTVDEALKDA QTRITK