Gene Spro_1287 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagSpro_1287 
Symbol 
ID5604873 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameSerratia proteamaculans 568 
KingdomBacteria 
Replicon accessionNC_009832 
Strand
Start bp1416323 
End bp1417378 
Gene Length1056 bp 
Protein Length351 aa 
Translation table11 
GC content56% 
IMG OID640936819 
Productphospho-2-dehydro-3-deoxyheptonate aldolase 
Protein accessionYP_001477519 
Protein GI157369530 
COG category[E] Amino acid transport and metabolism 
COG ID[COG0722] 3-deoxy-D-arabino-heptulosonate 7-phosphate (DAHP) synthase 
TIGRFAM ID[TIGR00034] phospho-2-dehydro-3-deoxyheptonate aldolase 


Plasmid Coverage information

Num covering plasmid clones24 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones13 
Fosmid unclonability p-value0.300084 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGAATTACC TAAATGACGA TTTACGTATT AAAGAAATTA AGGAACTCCT GCCACCGGTG 
GCCCTGTTAG AAAAGTTTCC GGCCACTGAA CGTGCAGCAG AGACGGTGTC ACAGGCACGC
AATGCCATTC ATCAGATCCT TCGCGGCAGC GACGATCGCC TGCTGGTGGT GATCGGCCCT
TGCTCGATCC ACGATACCAA AGCCGCCAAA GAGTACGCCG GTCGCCTGCT GGCGCTGCGT
CAGGAACTGA GCGGTGAGCT GGAAGTGGTG ATGCGCGTTT ATTTTGAAAA ACCGCGTACT
ACCGTGGGCT GGAAGGGTTT GATCAACGAT CCGCAGATGG ATAATAGCTT CCAGATCAAC
GACGGCCTGC GTTTGGCGCG TAAGCTGCTG CTGGATATCA ATGATTCTGG CCTGCCGGCT
GCCGGCGAAT TCCTCGACAT GATCACCCCG CAGTATCTGG CAGATCTGAT GAGCTGGGGC
GCGATCGGCG CACGCACCAC GGAATCTCAG GTGCACCGCG AACTGTCTTC CGGCCTGTCT
TGCCCGGTTG GCTTCAAAAA CGGCACCGAC GGTACCATTA AGGTGGCGAT TGATGCCATC
AACGCCGCCA GCGCGCCGCA CTGTTTCCTG TCGGTAACCA AATGGGGCCA CTCGGCCATT
GTTAACACCA GCGGTAACGG CGACTGTCAC ATCATTCTGC GCGGCGGCAA AGAGCCGAAC
TACAGCGCGG CGCACGTGAA ACAAGTCAAA GAAGGCCTGG TTAAAGCGGG TCTGCCTGCA
CAGGTGATGA TCGATTTTAG CCACGCCAAC AGCAGCAAGC AGTTCAAAAA GCAGCTGGAA
GTGAACGCAG ACGTCTGCCA ACAGATTGCC GGCGGTGAAA AGGCGATTAT GGGCGTGATG
ATCGAAAGCC ATCTGGTGGA AGGCAACCAG AACCTGGAAA GCGGCGATCC GCTGGTCTAC
GGCAAGAGCG TCACCGACGC CTGCATCGGC TGGTCAGACA CCGAAACTGT ACTGCGTGAA
CTGGCGGAAG CAGTGAAAGT GCGTCGCAAC AAGTAA
 
Protein sequence
MNYLNDDLRI KEIKELLPPV ALLEKFPATE RAAETVSQAR NAIHQILRGS DDRLLVVIGP 
CSIHDTKAAK EYAGRLLALR QELSGELEVV MRVYFEKPRT TVGWKGLIND PQMDNSFQIN
DGLRLARKLL LDINDSGLPA AGEFLDMITP QYLADLMSWG AIGARTTESQ VHRELSSGLS
CPVGFKNGTD GTIKVAIDAI NAASAPHCFL SVTKWGHSAI VNTSGNGDCH IILRGGKEPN
YSAAHVKQVK EGLVKAGLPA QVMIDFSHAN SSKQFKKQLE VNADVCQQIA GGEKAIMGVM
IESHLVEGNQ NLESGDPLVY GKSVTDACIG WSDTETVLRE LAEAVKVRRN K