Gene B21_03013 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagB21_03013 
SymbolkdsD 
ID8116653 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameEscherichia coli BL21 
KingdomBacteria 
Replicon accessionNC_012892 
Strand
Start bp3208973 
End bp3209959 
Gene Length987 bp 
Protein Length328 aa 
Translation table11 
GC content52% 
IMG OID644849198 
Producthypothetical protein 
Protein accessionYP_003000771 
Protein GI251786467 
COG category[M] Cell wall/membrane/envelope biogenesis 
COG ID[COG0794] Predicted sugar phosphate isomerase involved in capsule formation 
TIGRFAM ID[TIGR00393] KpsF/GutQ family protein 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.00296594 
Plasmid hitchhikingYes 
Plasmid clonabilityhitchhiker 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGTCGCACG TAGAGTTACA ACCGGGTTTT GACTTTCAGC AAGCAGGTAA AGAAGTCCTG 
GCGATTGAAC GTGAATGCCT GGCGGAGCTT GATCAATACA TCAATCAGAA TTTCACGCTT
GCCTGTGAAA AGATGTTCTG GTGTAAAGGG AAAGTTGTCG TCATGGGGAT GGGAAAATCG
GGGCATATTG GGCGAAAAAT GGCGGCAACG TTTGCCAGCA CCGGTACACC TTCATTTTTC
GTCCATCCTG GTGAAGCCGC GCATGGTGAT TTAGGCATGG TTACCCCACA GGATGTGGTG
ATTGCTATCT CTAACTCTGG TGAATCCAGC GAAATCACGG CCTTAATTCC AGTGCTTAAG
CGTCTTCACG TACCGTTAAT CTGCATCACC GGTCGCCCGG AGAGCAGCAT GGCGCGCGCC
GCAGATGTGC ATCTGTGTGT TAAAGTAGCG AAAGAAGCCT GTCCGTTAGG GCTGGCACCG
ACCAGCAGCA CCACCGCCAC GCTGGTTATG GGCGATGCCC TCGCTGTCGC GCTGTTAAAA
GCACGCGGCT TTACTGCTGA AGATTTTGCG CTCTCACACC CAGGCGGCGC ACTGGGTCGT
AAACTTCTGC TGCGCGTAAA CGATATTATG CATACGGGCG ATGAGATCCC GCATGTTAAG
AAAACGGCCA GTCTGCGTGA CGCATTGCTG GAAGTTACCC GCAAAAATCT TGGTATGACT
GTCATTTGCG ATGACAATAT GATGATTGAA GGCATCTTTA CCGACGGTGA TTTACGCCGT
GTCTTCGATA TGGGCGTGGA TGTTCGTCAG TTAAGTATTG CCGATGTGAT GACGCCGGGG
GGAATACGTG TGCGCCCTGG CATTCTGGCC GTTGAGGCAC TGAACTTAAT GCAGTCCCGC
CATATCACCT CCGTGATGGT TGCCGATGGC GACCATTTAC TCGGTGTGTT ACATATGCAT
GATTTACTGC GTGCAGGCGT AGTGTAA
 
Protein sequence
MSHVELQPGF DFQQAGKEVL AIERECLAEL DQYINQNFTL ACEKMFWCKG KVVVMGMGKS 
GHIGRKMAAT FASTGTPSFF VHPGEAAHGD LGMVTPQDVV IAISNSGESS EITALIPVLK
RLHVPLICIT GRPESSMARA ADVHLCVKVA KEACPLGLAP TSSTTATLVM GDALAVALLK
ARGFTAEDFA LSHPGGALGR KLLLRVNDIM HTGDEIPHVK KTASLRDALL EVTRKNLGMT
VICDDNMMIE GIFTDGDLRR VFDMGVDVRQ LSIADVMTPG GIRVRPGILA VEALNLMQSR
HITSVMVADG DHLLGVLHMH DLLRAGVV