Gene SbBS512_E4821 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagSbBS512_E4821 
SymbolpepA 
ID6271352 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameShigella boydii CDC 3083-94 
KingdomBacteria 
Replicon accessionNC_010658 
Strand
Start bp4492432 
End bp4493943 
Gene Length1512 bp 
Protein Length503 aa 
Translation table11 
GC content55% 
IMG OID641728562 
Productleucyl aminopeptidase 
Protein accessionYP_001882956 
Protein GI187730192 
COG category[E] Amino acid transport and metabolism 
COG ID[COG0260] Leucyl aminopeptidase 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.00000301619 
Plasmid hitchhikingYes 
Plasmid clonabilityhitchhiker 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGGAGTTTA GTGTAAAAAG CGGTAGCCCG GAGAAACAGC GGAGTGCCTG CATCGTCGTG 
GGCGTCTTCG AACCACGTCG CCTTTCTCCG ATTGCAGAAC AGCTCGATAA AATCAGCGAT
GGGTACATCA GCGCCCTGCT ACGTCGGGGC GAACTGGAAG GAAAACCGGG GCAGACATTG
TTGCTGCACC ATGTTCCGAA TGTACTTTCC GAGCGAATTC TCCTTATTGG TTGCGGCAAA
GAACGTGAGC TGGATGAGCG TCAGTACAAG CAGGTTATTC AGAAAACCAT TAATACCCTG
AATGATACTG GCTCAATGGA AGCGGTCTGC TTTCTGACTG AACTGCACGT TAAAGGCCGT
AACAACTACT GGAAAGTGCG TCAGGCTGTC GAGACGGCAA AAGAGACGCT CTACAGTTTC
GATCAGCTGA AAACGAACAA GAGCGAACCG CGTCGTCCGC TGCGTAAAAT GGTGTTCAAC
GTGCCGACCC GCCGTGAACT GACCAGCGGT GAGCGCGCGA TCCTGCACGG TCTGGCGATT
GCCGCCGCGA TTAAAGCAGC AAAAGATCTC GGCAATATGC CGCCGAATAT CTGTAACGCC
GCTTACCTCG CTTCACAAGC GCGCCAGCTG GCTGACAGCT ACAGCAAGAA TGTCATCACC
CGCGTTATCG GCGAACAGCA GATGAAAGAG CTGGGGATGC ATTCCTATCT GGCGGTCGGT
CAGGGTTCGC AAAACGAATC GCTGATGTCG GTGATTGAGT ACAAAGGCAA CGCGTCGGAA
GATGCACGCC CAATCGTGCT GGTGGGTAAA GGTTTAACCT TCGACTCCGG CGGTATCTCG
ATCAAGCCTT CAGAAGGCAT GGATGAGATG AAGTACGATA TGTGCGGTGC GGCAGCGGTT
TACGGCGTGA TGCGGATGGT CGCGGAGCTA CAACTGCCGA TTAACGTTAT CGGCGTGTTG
GCAGGCTGCG AAAACATGCC TGGCGGACGA GCCTATCGTC CGGGCGATGT GTTAACCACC
ATGTCCGGTC AAACTGTTGA AGTGCTGAAT ACCGACGCTG AAGGCCGCCT GGTACTGTGC
GACGTGTTAA CTTACGTTGA GCGTTTTGAG CCGGAAGCGG TGATTGACGT GGCGACGCTG
ACCGGTGCCT GCGTGATCGC GCTGGGTCAT CACATTACCG GTCTGATGGC GAACCATAAT
CCGCTGGCCC ATGAACTGAT TGCCGCGTCT GAACAATCCG GTGACCGCGC ATGGCGCTTA
CCGCTGGGTG ACGAGTATCA GGAACAGCTG GAGTCCAATT TTGCCGATAT GGCGAACATT
GGCGGTCGTC CTGGTGGGGC GATTACCGCA GGTTGCTTCC TGTCGCGCTT TACCCGTAAG
TACAACTGGG CGCACCTGGA TATCGCCGGA ACCGCCTGGC GTTCTGGTAA AGCAAAAGGC
GCCACCGGTC GTCCGGTGGC GTTGCTGGCA CAGTTCCTGT TAAACCGCGC TGGGTTTAAC
GGCGAAGAGT AA
 
Protein sequence
MEFSVKSGSP EKQRSACIVV GVFEPRRLSP IAEQLDKISD GYISALLRRG ELEGKPGQTL 
LLHHVPNVLS ERILLIGCGK ERELDERQYK QVIQKTINTL NDTGSMEAVC FLTELHVKGR
NNYWKVRQAV ETAKETLYSF DQLKTNKSEP RRPLRKMVFN VPTRRELTSG ERAILHGLAI
AAAIKAAKDL GNMPPNICNA AYLASQARQL ADSYSKNVIT RVIGEQQMKE LGMHSYLAVG
QGSQNESLMS VIEYKGNASE DARPIVLVGK GLTFDSGGIS IKPSEGMDEM KYDMCGAAAV
YGVMRMVAEL QLPINVIGVL AGCENMPGGR AYRPGDVLTT MSGQTVEVLN TDAEGRLVLC
DVLTYVERFE PEAVIDVATL TGACVIALGH HITGLMANHN PLAHELIAAS EQSGDRAWRL
PLGDEYQEQL ESNFADMANI GGRPGGAITA GCFLSRFTRK YNWAHLDIAG TAWRSGKAKG
ATGRPVALLA QFLLNRAGFN GEE