Gene CPF_2239 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagCPF_2239 
Symbol 
ID4202973 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameClostridium perfringens ATCC 13124 
KingdomBacteria 
Replicon accessionNC_008261 
Strand
Start bp2485767 
End bp2486924 
Gene Length1158 bp 
Protein Length385 aa 
Translation table11 
GC content33% 
IMG OID638083104 
Productputative amidohydrolase 
Protein accessionYP_696663 
Protein GI110800196 
COG category[Q] Secondary metabolites biosynthesis, transport and catabolism 
COG ID[COG1228] Imidazolonepropionase and related amidohydrolases 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.0161054 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGTTAATTA AAAATGGGAA AATATTTACC TGTGAAGAAG GTAAGATATA TGAAAAAGGT 
GATATTCTAA TTAAGGATGG AAAGATAAGT AGAATTGGGG AAGATTTAAG TCAATACATA
GGAGAAGAAG AGGTTATTGA TGCTAAAGGA CTATTAATAT TTCCAGGGTT TATTGAAGCA
CATTGTCATT TAGGACTACA TGAAGAAGGA AATAATGGGG CAGGAAATGG AACCAATGAA
GCTAGTGAGC CTATAACCCC ACAAATGAGA GCTATAGATG GAATAAATCC CTTTGATGGA
GGATTCCAAT CTGCAATGGA AGCAGGAGTT ACCACAGCTG TAATTGGGCC TGGAAGCGCT
AATGTAATAG GAGGACAGTT TGCCGCTGTA AAAACAAGTG GAATATGTAT TGATGACATG
ATAATAAAGG AACCTGTAGC AATAAAGGTT GCCTTTGGAG AAAATCCAAA AAGAGTTTAT
TCTGGAAAGA ATAAAATGCC TAATACAAGA ATGGCTATTG CAGCTTTATT AAGAGAAACT
TTAACAGAGG CTGTTAATTA TAAAAATAGA AAAATTGATG CTGAAATAGA GGATAGGGAT
TTTAGTAAGA ATTTAAAATA TGAGGCTTTA CTTCCCTTAA TTAATAGAGA AATACCTATG
AAAGCTCATA CCCATAGGGC AGATGATATT TTAACTGCCA TAAGAATAGC TAAGGAATTT
AATCTTAAAT TAACTTTAGA TCACTGTACA GAAGGACATT TAATAAGTGA TTATATTAAA
AGAGAAAACT TAGATGCTAT AGTTGGGCCA ACTTTAAGTT TTAATGGAAA GGCTGAGACT
TTAAATAAGA CCTTTAAGAC TCCAAAGGCC TTAATAGATA AAGGAATTAA AGTAGCAATA
ACTACAGACC ATCCAGTGGT AACAATAGAC AATCTTCCAC TTTGTGCAGC TATGGCTATG
AAAGAAGGAA TTACTTTTAA TGAGGCCTTA GAAGCAATAA CAATAAATCC AGCTGAAATA
ATAGGTATTG ATGAAAGGGT TGGAAGCTTA AAGGAAGGAA AGGATGGAGA TTTAGTAATT
TTAAATGGAA GTCCTTTTGA AATAGCTACA AAAACTATTT ATACAATTAT AAATGGAGAG
GTAGTTTATA AAGACTAG
 
Protein sequence
MLIKNGKIFT CEEGKIYEKG DILIKDGKIS RIGEDLSQYI GEEEVIDAKG LLIFPGFIEA 
HCHLGLHEEG NNGAGNGTNE ASEPITPQMR AIDGINPFDG GFQSAMEAGV TTAVIGPGSA
NVIGGQFAAV KTSGICIDDM IIKEPVAIKV AFGENPKRVY SGKNKMPNTR MAIAALLRET
LTEAVNYKNR KIDAEIEDRD FSKNLKYEAL LPLINREIPM KAHTHRADDI LTAIRIAKEF
NLKLTLDHCT EGHLISDYIK RENLDAIVGP TLSFNGKAET LNKTFKTPKA LIDKGIKVAI
TTDHPVVTID NLPLCAAMAM KEGITFNEAL EAITINPAEI IGIDERVGSL KEGKDGDLVI
LNGSPFEIAT KTIYTIINGE VVYKD