Gene B21_02255 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagB21_02255 
SymbolypdE 
ID8115837 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameEscherichia coli BL21 
KingdomBacteria 
Replicon accessionNC_012892 
Strand
Start bp2381705 
End bp2382742 
Gene Length1038 bp 
Protein Length345 aa 
Translation table11 
GC content58% 
IMG OID644848459 
Producthypothetical protein 
Protein accessionYP_003000032 
Protein GI251785728 
COG category[G] Carbohydrate transport and metabolism 
COG ID[COG1363] Cellulase M and related proteins 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones26 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGGATTTAT CGCTATTAAA AGCGTTGAGC GAGGCAGATG CGATCGCCTC CTCGGAACAG 
GAAGTGCGGC AGATCCTGCT GGAAGAAGCG GATCGCCTGC AAAAAGAAGT GCGATTTGAT
GGTCTGGGAT CGGTGCTGAT CCGCCTGAAT GAATCGACAG GTCCGAAGGT GATGATCTGT
GCGCATATGG ACGAAGTGGG ATTTATGGTG CACAGCATCT CCCGCGAAGG GGCGATTGAT
GTGCTGCCGG TTGGCAACGT ACGCATGGCT GCCCGCCAGC TGCAGCCGGT GCGCATCACC
ACCCGTGAAG AGTGCAAAAT TCCAGGCCTG CTTGACGGCG ACCGGCAGGG GAATGACGTC
AGCGCCATGC GCGTGGACAT TGGTGCGCGC TCCTATGACG AAGTGATGCA GGCGGGAATT
CGTCCCGGCG ATCGCGTCAC GTTTGATACC ACTTTTCAGG TTCTCCCTCA CCAGCGAGTG
ATGGGGAAAG CCTTTGATGA CCGCCTCGGT TGCTATCTGC TGGTGACGTT ACTGCGCGAA
CTGCACGACG CCGAACTACC TGCGGAAGTG TGGCTGGTGG CAAGTTCCAG CGAAGAGGTG
GGATTACGCG GCGGGCAAAC TGCCACCCGC GCGGTGTCGC CGGACGTCGC CATTGTGCTT
GATACCGCCT GCTGGGCGAA AAACTTTGAT TATGGCGCGG CTAACCATCG CCAGATTGGT
AACGGGCCGA TGCTGGTGTT AAGCGACAAG TCGCTGATTG CGCCGCCAAA ACTTACCGCC
TGGGTCGAAA CCGTGGCGGC AGAAATTGGC GTGCCGTTGC AGGCAGATAT GTTCAGCAAC
GGCGGCACGG ATGGCGGGGC GGTGCACTTA ACCGGCACCG GCGTGCCCAC AGTGGTGATG
GGGCCAGCAA CCCGCCATGG ACATTGCGCC GCATCGATTG CCGATTGCCG CGACATTTTG
CAGATGCAGC AACTTTTATC TGCCCTTATT CAACGTCTTA CGCGTGAGAC GGTTGTTCAA
CTGACGGATT TCAGATGA
 
Protein sequence
MDLSLLKALS EADAIASSEQ EVRQILLEEA DRLQKEVRFD GLGSVLIRLN ESTGPKVMIC 
AHMDEVGFMV HSISREGAID VLPVGNVRMA ARQLQPVRIT TREECKIPGL LDGDRQGNDV
SAMRVDIGAR SYDEVMQAGI RPGDRVTFDT TFQVLPHQRV MGKAFDDRLG CYLLVTLLRE
LHDAELPAEV WLVASSSEEV GLRGGQTATR AVSPDVAIVL DTACWAKNFD YGAANHRQIG
NGPMLVLSDK SLIAPPKLTA WVETVAAEIG VPLQADMFSN GGTDGGAVHL TGTGVPTVVM
GPATRHGHCA ASIADCRDIL QMQQLLSALI QRLTRETVVQ LTDFR