Gene SbBS512_E0986 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagSbBS512_E0986 
Symbolamn 
ID6272679 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameShigella boydii CDC 3083-94 
KingdomBacteria 
Replicon accessionNC_010658 
Strand
Start bp906987 
End bp908135 
Gene Length1149 bp 
Protein Length382 aa 
Translation table11 
GC content50% 
IMG OID641725134 
ProductAMP nucleosidase 
Protein accessionYP_001879658 
Protein GI187733016 
COG category[F] Nucleotide transport and metabolism 
COG ID[COG0775] Nucleoside phosphorylase 
TIGRFAM ID[TIGR01717] AMP nucleosidase 


Plasmid Coverage information

Num covering plasmid clones10 
Plasmid unclonability p-value0.00000563719 
Plasmid hitchhikingYes 
Plasmid clonabilityhitchhiker 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
TTGCTGTATC AGGATTATGG TGCGCATATC TCAGTGCAAC CCTCGCAGCA TGAAATCCCT 
TATCCTTATG TCATCGATGG CTCTGAATTG ACACTTGATC GCTCAATGAG CGCTGGGTTA
ACTCGCTACT TTCCGACAAC AGAACTGGCG CAAATTGGCG ATGAAACTGC AGACGGCATT
TATCATCCAA CTGAATTCTC CCCGCTATCG CATTTTGATG CGCGCCGCGT CGATTTTTCC
CTCGCACGGT TGCGCCATTA TACCGGTACG CCAGTTGAAC ATTTTCAGCC GTTCGTCTTG
TTTACCAACT ACACACGTTA TGTGGATGAA TTCGTTCGTT GGGGATGCAG CCAGATCCTC
GATCCTGATA GTCCCTACAT TGCCCTTTCT TGTGCTGGCG GGAACTGGAT CACCGCCGAA
ACCGAAGCGC CAGAAGAAGC CATTTCCGAC CTTGCATGGA AAAAACATCA GATGCCAGCA
TGGCATTTAA TTACCGCCGA TGGTCAGGGT ATTACTCTGG TGAATATTGG CATAGGACCG
TCAAATGCTA AAACCATCTG CGATCATCTG GCAGTGCTAC GCCCGGATGT CTGGTTGATG
ATTGGTCACT GTGGCGGATT ACGTGAAAGT CAGGCCATTG GCGATTATGT ACTTGCACAC
GCTTATTTAC GCGATGACCA CGTTCTTGAT GCGGTTCTGC CGCCCGATAT TCCTATTCCG
AGCATTGCTG AAGTGCAACG TGCGCTTTAT GACGCCACCA AGCTGGTGAG TGGCAGGCCC
GGTGAGGAAG TCAAACAGCG GCTACGTACT GGTACTGTGG TAACCACAGA TGACAGGAAC
TGGGAATTAC GTTACTCAGC TTCTGCACTT CGTTTTAACT TAAGCCGGGC CGTAGCAATT
GATATGGAAA GTGCAACCAT TGCCGCGCAA GGATATCGTT TCCGCGTGCC ATACGGGACA
CTACTGTGTG TTTCAGATAA ACCGTTGCAT GGCGAGATTA AACTTCCCGG TCAGGCTAAC
CGTTTTTATG AAGGCGCTAT TTCCGAACAC CTACAAATTG GCATTCGGGC GATCGATTTG
CTGCGCGCAG AAGGCGACCG ACTGCATTCA CGTAAATTAC GAACCTTTAA TGAGCCGCCG
TTCCGATAA
 
Protein sequence
MLYQDYGAHI SVQPSQHEIP YPYVIDGSEL TLDRSMSAGL TRYFPTTELA QIGDETADGI 
YHPTEFSPLS HFDARRVDFS LARLRHYTGT PVEHFQPFVL FTNYTRYVDE FVRWGCSQIL
DPDSPYIALS CAGGNWITAE TEAPEEAISD LAWKKHQMPA WHLITADGQG ITLVNIGIGP
SNAKTICDHL AVLRPDVWLM IGHCGGLRES QAIGDYVLAH AYLRDDHVLD AVLPPDIPIP
SIAEVQRALY DATKLVSGRP GEEVKQRLRT GTVVTTDDRN WELRYSASAL RFNLSRAVAI
DMESATIAAQ GYRFRVPYGT LLCVSDKPLH GEIKLPGQAN RFYEGAISEH LQIGIRAIDL
LRAEGDRLHS RKLRTFNEPP FR