Gene EcE24377A_3685 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagEcE24377A_3685 
SymbolkdsD 
ID5587199 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameEscherichia coli E24377A 
KingdomBacteria 
Replicon accessionNC_009801 
Strand
Start bp3678423 
End bp3679409 
Gene Length987 bp 
Protein Length328 aa 
Translation table11 
GC content52% 
IMG OID640927308 
ProductD-arabinose 5-phosphate isomerase 
Protein accessionYP_001464675 
Protein GI157158811 
COG category[M] Cell wall/membrane/envelope biogenesis
[T] Signal transduction mechanisms 
COG ID[COG0794] Predicted sugar phosphate isomerase involved in capsule formation
[COG2905] Predicted signal-transduction protein containing cAMP-binding and CBS domains 
TIGRFAM ID[TIGR00393] KpsF/GutQ family protein 


Plasmid Coverage information

Num covering plasmid clones29 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGTCGCACG TAGAGTTACA ACCGGGTTTT GACTTTCAGC AAGCAGGTAA AGAAGTCCTG 
GCGATTGAAC GTGAATGCCT GGCGGAGCTT GATCAATACA TCAATCAGAA TTTCACGCTT
GCCTGTGAAA AGATGTTCTG GTGTAAAGGG AAAGTTGTCG TCATGGGGAT GGGGAAATCG
GGGCACATCG GGCGCAAAAT GGCCGCAACG TTTGCCAGCA CCGGTACACC TTCATTTTTC
GTCCATCCTG GTGAAGCCGC GCATGGTGAT TTAGGCATGG TCACCCCACA GGATGTGGTG
ATTGCTATCT CTAACTCAGG TGAATCCAGC GAAATCACGG CCTTAATTCC AGTGCTTAAG
CGTCTTCACG TACCGTTAAT CTGCATCACC GGTCGCCCGG AGAGCAGCAT GGCGCGCGCC
GCAGATGTGC ATCTGTGTGT TAAAGTAGCG AAAGAAGCCT GTCCGTTAGG GCTGGCACCG
ACCAGCAGCA CCACCGCCAC GCTGGTTATG GGCGATGCCC TCGCTGTCGC GCTGTTAAAA
GCACGCGGCT TTACTGCTGA AGATTTTGCG CTCTCACACC CAGGCGGCGC ACTGGGTCGT
AAACTTCTGC TGCGCGTAAA CGATATTATG CATACGGGCG ATGAGATCCC GCATGTTAAG
AAAACGGCCA GTCTGCGTGA CGCATTGCTG GAAGTTACCC GCAAAAATCT TGGTATGACT
GTCATTTGCG ATGACAATAT GATGATTGAA GGCATCTTTA CCGACGGTGA TTTACGCCGT
GTCTTCGATA TGGGCGTGGA TGTTCGTCAG TTAAGTATTG CCGATGTGAT GACGCCGGGG
GGAATACGTG TGCGCCCTGG CATTCTGGCC GTTGAGGCAC TGAACTTAAT GCAGTCCCGC
CATATCACCT CCGTGATGGT TGCCGATGGC GACCATTTAC TCGGTGTGTT ACATATGCAT
GATTTACTGC GTGCAGGCGT AGTGTAA
 
Protein sequence
MSHVELQPGF DFQQAGKEVL AIERECLAEL DQYINQNFTL ACEKMFWCKG KVVVMGMGKS 
GHIGRKMAAT FASTGTPSFF VHPGEAAHGD LGMVTPQDVV IAISNSGESS EITALIPVLK
RLHVPLICIT GRPESSMARA ADVHLCVKVA KEACPLGLAP TSSTTATLVM GDALAVALLK
ARGFTAEDFA LSHPGGALGR KLLLRVNDIM HTGDEIPHVK KTASLRDALL EVTRKNLGMT
VICDDNMMIE GIFTDGDLRR VFDMGVDVRQ LSIADVMTPG GIRVRPGILA VEALNLMQSR
HITSVMVADG DHLLGVLHMH DLLRAGVV