Gene EcHS_A3390 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagEcHS_A3390 
SymbolkdsD 
ID5593418 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameEscherichia coli HS 
KingdomBacteria 
Replicon accessionNC_009800 
Strand
Start bp3392193 
End bp3393179 
Gene Length987 bp 
Protein Length328 aa 
Translation table11 
GC content52% 
IMG OID640922510 
ProductD-arabinose 5-phosphate isomerase 
Protein accessionYP_001459999 
Protein GI157162681 
COG category[M] Cell wall/membrane/envelope biogenesis
[T] Signal transduction mechanisms 
COG ID[COG0794] Predicted sugar phosphate isomerase involved in capsule formation
[COG2905] Predicted signal-transduction protein containing cAMP-binding and CBS domains 
TIGRFAM ID[TIGR00393] KpsF/GutQ family protein 


Plasmid Coverage information

Num covering plasmid clones45 
Plasmid unclonability p-value0.997652 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGTCGCACG TAGAGTTACA ACCGGGTTTT GACTTTCAGC AAGCAGGTAA AGAAGTCCTG 
GCGATTGAAC GTGAATGCCT GGCGGAGCTT GATCAATACA TCAATCAGAA TTTCACGCTT
GCCTGTGAAA AGATGTTCTG GTGTAAAGGG AAAGTTGTCG TCATGGGGAT GGGAAAATCG
GGGCATATTG GGCGAAAAAT GGCGGCAACG TTTGCCAGCA CCGGTACACC TTCATTTTTC
GTCCATCCTG GTGAAGCCGC GCATGGTGAT TTAGGCATGG TTACCCCACA GGATGTGGTG
ATTGCTATCT CTAACTCTGG TGAATCCAGC GAAATCACGG CCTTAATTCC AGTGCTTAAG
CGTCTTCACG TACCGTTAAT CTGCATCACC GGTCGCCCGG AGAGCAGCAT GGCGCGCGCC
GCAGATGTGC ATCTGTGTGT TAAAGTAGCG AAAGAAGCCT GTCCGTTAGG GCTGGCACCG
ACCAGCAGCA CCACCGCCAC GCTGGTTATG GGCGATGCCC TCGCTGTCGC GCTGTTAAAA
GCACGCGGCT TTACTGCTGA AGATTTTGCG CTCTCACACC CAGGCGGCGC ACTGGGTCGT
AAACTTCTGC TGCGCGTAAA CGATATTATG CATACGGGCG ATGAGATCCC GCATGTTAAG
AAAACGGCCA GTCTGCGTGA CGCATTGCTG GAAGTTACCC GCAAAAATCT TGGTATGACT
GTCATTTGCG ATGACAATAT GATGATTGAA GGCATCTTTA CCGACGGTGA TTTACGCCGT
GTCTTCGATA TGGGCGTGGA TGTTCGTCAG TTAAGTATTG CCGATGTGAT GACGCCGGGG
GGAATACGTG TGCGCCCTGG CATTCTGGCC GTTGAGGCAC TGAACTTAAT GCAGTCCCGC
CATATCACCT CCGTGATGGT TGCCGATGGC GACCATTTAC TCGGTGTGTT ACATATGCAT
GATTTACTGC GTGCAGGCGT AGTGTAA
 
Protein sequence
MSHVELQPGF DFQQAGKEVL AIERECLAEL DQYINQNFTL ACEKMFWCKG KVVVMGMGKS 
GHIGRKMAAT FASTGTPSFF VHPGEAAHGD LGMVTPQDVV IAISNSGESS EITALIPVLK
RLHVPLICIT GRPESSMARA ADVHLCVKVA KEACPLGLAP TSSTTATLVM GDALAVALLK
ARGFTAEDFA LSHPGGALGR KLLLRVNDIM HTGDEIPHVK KTASLRDALL EVTRKNLGMT
VICDDNMMIE GIFTDGDLRR VFDMGVDVRQ LSIADVMTPG GIRVRPGILA VEALNLMQSR
HITSVMVADG DHLLGVLHMH DLLRAGVV