Gene ECH74115_1076 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagECH74115_1076 
SymbollpxK 
ID6969538 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameEscherichia coli O157:H7 str. EC4115 
KingdomBacteria 
Replicon accessionNC_011353 
Strand
Start bp1101364 
End bp1102350 
Gene Length987 bp 
Protein Length328 aa 
Translation table11 
GC content55% 
IMG OID643385088 
Producttetraacyldisaccharide 4'-kinase 
Protein accessionYP_002269587 
Protein GI209399961 
COG category[M] Cell wall/membrane/envelope biogenesis 
COG ID[COG1663] Tetraacyldisaccharide-1-P 4'-kinase 
TIGRFAM ID[TIGR00682] tetraacyldisaccharide 4'-kinase 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.0122696 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones64 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGATCGAAA AAATCTGGTC TGGTGAATCC CCTTTGTGGC GGCTATTGCT GCCACTCTCC 
TGGTTGTATG GCCTGGTGAG TGGCGCGATC CGTCTTTGCT ATAAACTAAA ACTGAAGCGC
GCCTGGCGTG CCCCCGTACC GGTTGTCGTG GTTGGTAATC TCACCGCAGG CGGCAACGGA
AAAACCCCGG TCGTTGTCTG GCTGGTGGAA CAGTTGCAAC AGCGCGGTAT TCGCGTGGGG
GTCGTATCGC GGGGATATGG TGGTAAGGCT GAATCTTATC CGCTGTTATT GTCGGCAGAT
ACCACCACAG CACAGGCGGG TGATGAACCT GTGTTGATTT ATCAACGCAC TGATGCGCCT
GTTGCGGTTT CTCCCGTTCG TTCTGATGCG GTAAAAGCCA TTCTGGCGCA ACACCCTGAT
GTGCAGATCA TCGTAACCGA CGACGGTTTA CAGCATTACC GTCTGGCGCG TGATGTGGAA
ATTGTCGTTA TTGATGGTGT GCGTCGCTTT GGCAATGGCT GGTGGTTGCC GGCGGGGCCA
ATGCGTGAGC GAGCGGGGCG CTTAAAGTCA GTTGATGCGG TAATCGTCAA CGGCGGTGTC
CCCCGCAGCG GTGAAATCCC CATGCATCTG CTGCCGGGTC AGGCGGTGAA TTTACGTACC
GGTACGCGTT GTGACGTTGC TCAGCTTGAA CATGTGGTGG CGATGGCAGG GATTGGGCAT
CCGCCGCGCT TTTTTGCCAC GCTGAAGATG TGCGGCGTAC AACCGGAAAA ATGTGTACCG
CTGGCCGATC ATCAGTCTTT GAACCATGCG GATGTCAGCG CGTTGGTAAG CACCGGGCAA
ACGCTGGTAA TGACTGAAAA AGATGCGGTG AAATGCCGGG CCTTTGCAGA AGAAAATTGG
TGGTATTTGC CCGTTGACGC ACAGCTTTCA GGTGATGAAC CAGCGAAACT GCTTGCGCAA
CTAACCTCGC TGGCTTCTGG CAACTAG
 
Protein sequence
MIEKIWSGES PLWRLLLPLS WLYGLVSGAI RLCYKLKLKR AWRAPVPVVV VGNLTAGGNG 
KTPVVVWLVE QLQQRGIRVG VVSRGYGGKA ESYPLLLSAD TTTAQAGDEP VLIYQRTDAP
VAVSPVRSDA VKAILAQHPD VQIIVTDDGL QHYRLARDVE IVVIDGVRRF GNGWWLPAGP
MRERAGRLKS VDAVIVNGGV PRSGEIPMHL LPGQAVNLRT GTRCDVAQLE HVVAMAGIGH
PPRFFATLKM CGVQPEKCVP LADHQSLNHA DVSALVSTGQ TLVMTEKDAV KCRAFAEENW
WYLPVDAQLS GDEPAKLLAQ LTSLASGN