Gene SeSA_A4191 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagSeSA_A4191 
SymbolpepQ 
ID6516298 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameSalmonella enterica subsp. enterica serovar Schwarzengrund str. CVM19633 
KingdomBacteria 
Replicon accessionNC_011094 
Strand
Start bp4070071 
End bp4071402 
Gene Length1332 bp 
Protein Length443 aa 
Translation table11 
GC content53% 
IMG OID642749157 
Productproline dipeptidase 
Protein accessionYP_002116909 
Protein GI194734586 
COG category[E] Amino acid transport and metabolism 
COG ID[COG0006] Xaa-Pro aminopeptidase 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones12 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones20 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGGAATCAC TGGCCGCGCT CTATAAAAAT CATATTGTTA CCTTACAAGA ACGGACGCGC 
GATGTTCTGG CGCGCTTTAA GCTGGATGCG TTACTTATTC ATTCCGGCGA GCTTTTCAAC
GTCTTTCTCG ACGATCACCC TTATCCGTTT AAGGTCAATC CACAGTTTAA AGCGTGGGTG
CCGGTAACTC AGGTTCCAAA TTGCTGGCTG CTGGTCGATG GCGTCAACAA ACCCAAATTG
TGGTTTTATC TGCCGGTCGA TTACTGGCAT AACGTTGAAC CGCTGCCAAC GTCCTTCTGG
ACAGAAGAAG TCGAGGTCGT CGCCTTACCG AAAGCGGATG GCATCGGCAG CCAACTGCCT
GCCGCGCGTG GCAATATCGG CTATATCGGC CCGGTTCCTG AGCGCGCGCT ACAATTGGAT
ATCGCTGCCA GCAACATCAA CCCGAAAGGT GTTATCGACT ATCTGCATTA CTACCGCGCC
TATAAAACGG ATTATGAACT GGCCTGTATG CGCGAAGCGC AGAAAATGGC GGTGAGCGGT
CATCGGGCGG CGGAAGAGGC CTTCCGTTCC GGCATGAGCG AGTTTGACAT CAACCTGGCG
TACCTGACCG CCACGGGACA TCGCGATACC GATGTTCCGT ACAGCAACAT TGTGGCGCTG
AACGAACATG CCGCCGTGCT GCATTACACG AAACTGGATC ATCAGGCACC GTCTGAAATG
CGCAGTTTCC TGCTGGATGC GGGCGCGGAA TACAACGGCT ACGCGGCGGA TCTGACGCGA
ACCTGGTCGG CGAAAAGCGA TAACGACTAC GCCCACTTGG TGAAAGATGT TAACGACGAA
CAGTTGGCGC TGATCGCTAC CATGAAGGCG GGCGTCAGCT ATGTGGATTA TCATATTCAG
TTCCATCAAC GCATCGCGAA GCTGCTGCGT AAACATCAAA TCATTACCGA CATGAGTGAA
GAGGCGATGG TGGAAAATGA TCTCACCGGG CCGTTTATGC CGCACGGTAT TGGTCATCCG
TTGGGTCTGC AGGTACACGA TGTGGCCGGG TTTATGCAAG ATGATTCCGG TACGCATCTC
GCCGCGCCGT CCAAATACCC GTATCTGCGC TGCACGCGTG TGTTACAGCC GCGAATGGTG
TTGACCATCG AACCGGGGAT TTACTTCATC GAATCGCTGT TAGCGCCGTG GCGCGAAGGG
CCATTCAGCA AGCACTTCAA CTGGCAGAAA ATTGAAGCGC TCAAGCCTTT CGGCGGTATT
CGCATTGAAG ATAACGTGGT CATCCACGAA AACGGCGTGG AAAACATGAC GCGGGATTTA
AAACTGGCGT AA
 
Protein sequence
MESLAALYKN HIVTLQERTR DVLARFKLDA LLIHSGELFN VFLDDHPYPF KVNPQFKAWV 
PVTQVPNCWL LVDGVNKPKL WFYLPVDYWH NVEPLPTSFW TEEVEVVALP KADGIGSQLP
AARGNIGYIG PVPERALQLD IAASNINPKG VIDYLHYYRA YKTDYELACM REAQKMAVSG
HRAAEEAFRS GMSEFDINLA YLTATGHRDT DVPYSNIVAL NEHAAVLHYT KLDHQAPSEM
RSFLLDAGAE YNGYAADLTR TWSAKSDNDY AHLVKDVNDE QLALIATMKA GVSYVDYHIQ
FHQRIAKLLR KHQIITDMSE EAMVENDLTG PFMPHGIGHP LGLQVHDVAG FMQDDSGTHL
AAPSKYPYLR CTRVLQPRMV LTIEPGIYFI ESLLAPWREG PFSKHFNWQK IEALKPFGGI
RIEDNVVIHE NGVENMTRDL KLA