Gene EcolC_0503 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagEcolC_0503 
Symbol 
ID6068624 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameEscherichia coli ATCC 8739 
KingdomBacteria 
Replicon accessionNC_010468 
Strand
Start bp546107 
End bp547093 
Gene Length987 bp 
Protein Length328 aa 
Translation table11 
GC content52% 
IMG OID641599908 
ProductD-arabinose 5-phosphate isomerase 
Protein accessionYP_001723507 
Protein GI170018553 
COG category[M] Cell wall/membrane/envelope biogenesis
[T] Signal transduction mechanisms 
COG ID[COG0794] Predicted sugar phosphate isomerase involved in capsule formation
[COG2905] Predicted signal-transduction protein containing cAMP-binding and CBS domains 
TIGRFAM ID[TIGR00393] KpsF/GutQ family protein 


Plasmid Coverage information

Num covering plasmid clones19 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones13 
Fosmid unclonability p-value0.183909 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGTCGCACG TAGAGTTACA ACCGGGTTTT GACTTTCAGC AAGCAGGTAA AGAAGTCCTG 
GCGATTGAAC GTGAATGCCT GGCGGAGCTT GATCAATACA TCAATCAGAA TTTCACGCTT
GCCTGTGAAA AGATGTTCTG GTGTAAAGGG AAAGTTGTCG TCATGGGGAT GGGAAAATCG
GGGCATATTG GGCGAAAAAT GGCGGCAACG TTTGCCAGCA CCGGTACACC TTCATTTTTC
GTCCATCCTG GTGAAGCCGC GCATGGTGAT TTAGGCATGG TTACCCCACA GGATGTGGTG
ATTGCTATCT CTAACTCTGG TGAATCCAGC GAAATCACGG CCTTAATTCC AGTGCTTAAG
CGTCTTCACG TACCGTTAAT CTGCATCACC GGTCGCCCGG AGAGCAGCAT GGCGCGCGCC
GCAGATGTGC ATCTGTGTGT TAAAGTAGCG AAAGAAGCCT GTCCGTTAGG GCTGGCACCG
ACCAGCAGCA CCACCGCCAC GCTGGTTATG GGCGATGCCC TCGCTGTCGC GCTGTTAAAA
GCACGCGGCT TTACTGCTGA AGATTTTGCG CTCTCACACC CAGGCGGCGC ACTGGGTCGT
AAACTTCTGC TGCGCGTAAA CGATATTATG CATACGGGCG ATGAGATCCC GCATGTTAAG
AAAACGGCCA GTCTGCGTGA CGCATTGCTG GAAGTTACCC GCAAAAATCT TGGTATGACT
GTCATTTGCG ATGACAATAT GATGATTGAA GGCATCTTTA CCGACGGTGA TTTACGCCGT
GTCTTCGATA TGGGCGTGGA TGTTCGTCAG TTAAGTATTG CCGATGTGAT GACGCCGGGG
GGAATACGTG TGCGCCCTGG CATTCTGGCC GTTGAGGCAC TGAACTTAAT GCAGTCCCGC
CATATCACCT CCGTGATGGT TGCCGATGGC GACCATTTAC TCGGTGTGTT ACATATGCAT
GATTTACTGC GTGCAGGCGT AGTGTAA
 
Protein sequence
MSHVELQPGF DFQQAGKEVL AIERECLAEL DQYINQNFTL ACEKMFWCKG KVVVMGMGKS 
GHIGRKMAAT FASTGTPSFF VHPGEAAHGD LGMVTPQDVV IAISNSGESS EITALIPVLK
RLHVPLICIT GRPESSMARA ADVHLCVKVA KEACPLGLAP TSSTTATLVM GDALAVALLK
ARGFTAEDFA LSHPGGALGR KLLLRVNDIM HTGDEIPHVK KTASLRDALL EVTRKNLGMT
VICDDNMMIE GIFTDGDLRR VFDMGVDVRQ LSIADVMTPG GIRVRPGILA VEALNLMQSR
HITSVMVADG DHLLGVLHMH DLLRAGVV