Gene CPF_0605 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagCPF_0605 
Symbol 
ID4203264 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameClostridium perfringens ATCC 13124 
KingdomBacteria 
Replicon accessionNC_008261 
Strand
Start bp727443 
End bp729866 
Gene Length2424 bp 
Protein Length807 aa 
Translation table11 
GC content29% 
IMG OID638081490 
Productcell wall binding repeat-containing protein/mannosyl-glycoprotein endo-beta-N-acetylglucosamidase domain-containing protein 
Protein accessionYP_695058 
Protein GI110799389 
COG category[R] General function prediction only 
COG ID[COG5263] FOG: Glucan-binding domain (YG repeat) 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.187255 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
TTGATTAAAA AGTTTATTCT TGCTACAGTT ATTACTTTAA GTGTAACTTC AATTAATGTT 
CTAGCTTCAG AGAATGTAAA TGATAAGTCT GATGAAAATA AGTCTTCTAC TTTAGTTGGA
GAAACTTATG AAATTGTAAA AGAGCCAAAT ATGTTTTTTT CAATTCCAGG TGATAATGTA
GAAAGCTATA AAGATGAAAA TGGAATTGAG AGAGAAGAAT TTAAAAAGGA TAAAGAAGGT
CAAGAATCTA TAAACAGACT TGTTACAAAG ACTAAGAGTA AGTATGAGAT AGCATTAGCT
CATGAAAATG GGAAATATAC ATTTTTAGAT TCTGCAAACA CAAAGGAAGA AGCAGAGAAA
AAGGTTGAAA ATGCAAGTGA AAAATATAAT ACTTTTGCAG CAATGCCTGT TGTTTTAAAT
GATAGTGGAC AAGTAGCTTA TTCAGAAAAA TCAATGGGAA GATTAGTTAA ATATAAAAAT
GGAAGTCCAG CTGGATATGG AGAAATAACT AATATATATG CTAATCCTAA TTTAACAAAT
GACTTTACAT ATATTAATCA TGGTTATGTA GATGATGTTC CAATAATTGA GGATAGAGGA
AATGTTGCAA AAATAGAAGT TGGAGGATAT GAAGGTTGGG TAAATAAAGA TACTAGTTCT
GGAAATTATG ATTTAGTCAT AGTACCATTA AACCAAGTAA AAAATCCAAG TTATTATATA
GTTAGAGATG GAGAGTTAAT CCATTATATA AGTAGTGATC TTACCAATTA TTCTGAAGGT
GGATATGAGG TTGTTATAGG ACCAGCACCT AATTTCTTAA GTGAAAATGT AAAATATTAT
AGTTATGATA ACAAATATTT TTATAAAGAT TTAAGTACAT TAATAGGTGA TTTACAAAAT
GATAATCATA ATAATTCCGT AAATGCAAAT AATCCTTTTT ATCCATATTA TATGCATTTA
CCATTTAGAA GTAAAACAAC ATTTACTGCT GAAGAATTAA ATAATTTTAT TGCTAAAAAG
ACAAAATCAT ATAGCAAATT AAGAGGAACT GGACAAGCTT TTATAGATGC ACAAAATAAA
TATGGAGCAA ATGCTCTTTT ACTTTTAGGA CTAGCGGCTA ATGAATCAGC TTGGGGAACA
TCTCAAATAG CTCAACAAAA AAATAATTTA TTTGGAATAA ATGCTATTGA TTCAAGTCCA
GGAGCATCAG CAAATTCATT TGAAACTGTT GAAGGCTGTA TAAATGATTT TGCAAAATAT
TATATTTCTA GAGGATATTC AGATCCAGAA GATTGGAGAT ATTTTGGAGG ATACCTTGGA
AATAAGGGTA GTGGAGCTAA TGTTAAATAT GCTTCAGATC CATTCTGGGG AGAAAAAGCA
GGACAAAATG CTTATATAGC TGATTACTGG ACAAGTGGAA AAGGAATAGC AGGTTTAAAA
GATTATAATT ACTATCAGCT AGGCATTTAT ACAGGGGCTA GTAGTGTTAC AAATAAAGAT
AATGAAAAAT TATATGATGT AGGTAGCTTA TATACTGAAA GAGTAGAAAA AATAGGAGCT
ACAACAATTT TAACTAGCAA AGAAAAGATT AATCACAATG GAAAAGAATG TTATGAAATA
AATCCAGTTA GAACAACTCC TGTAATATCA AATGGATCAC CAGTAGCATT CCCAGGTCCA
TACGATTGGA ATGATAAAGG CTTTGTAGAT GCATCAAAAG TTAAGCTTAT AAATGAAGGT
AAGTATTCTG AAAATGTTAT GGGAAAATGG GAATTAAATA ATGGCATATG GTACTATTAT
ATAGACGGAA AATATGTTAC TGGGTGGAAT AATATAGATA ATAACAAGTA TTATTTTGAT
CAATCAGGTA AGATGCAAAC AGGTTGGCTA TGTATAGATA ATATATGGTA TCATTTTAGA
GATAATGGAA CAATAGACAT AGGTTGGAGA GAAATAGATG GATATTGGTA TTATTTTAAT
AATGATGGAG AAATGCAAAA GGGTTGGCAG ACAATAGGTA AATATAAGTA TTATTTCAAT
GATAATGGAG TAATGAATAT TGGCTGGAGA AAAATAGATA ATATTTGGTA CTATTTTAAT
ATAAATGGAG AAATGCAAAC AGGATGGTTA TTTGAAAATA ATATATGGTA TCATTTTGAA
ACTACAGGTT CAATGACTGT AGGATGGAAA AAAATAAATA ATGATTGGTA CTATTTTAAT
TCAAATGGAG AAATGCAAAC AGGATGGCAA AATATTGGTA CTCCAAAATA TCATTTTAGA
GATAATGGTG TAATGGACAT TGGATGGAGA GAAATAGATG GACAATGGCA TTATTTTAAT
GAGAATGGAG AAATGCAAAC AGGTTGGCAA TATATAAATG GAAATAGTTA TTATTTCGAT
GAAAAAGGAA TATGGAAACA GTAA
 
Protein sequence
MIKKFILATV ITLSVTSINV LASENVNDKS DENKSSTLVG ETYEIVKEPN MFFSIPGDNV 
ESYKDENGIE REEFKKDKEG QESINRLVTK TKSKYEIALA HENGKYTFLD SANTKEEAEK
KVENASEKYN TFAAMPVVLN DSGQVAYSEK SMGRLVKYKN GSPAGYGEIT NIYANPNLTN
DFTYINHGYV DDVPIIEDRG NVAKIEVGGY EGWVNKDTSS GNYDLVIVPL NQVKNPSYYI
VRDGELIHYI SSDLTNYSEG GYEVVIGPAP NFLSENVKYY SYDNKYFYKD LSTLIGDLQN
DNHNNSVNAN NPFYPYYMHL PFRSKTTFTA EELNNFIAKK TKSYSKLRGT GQAFIDAQNK
YGANALLLLG LAANESAWGT SQIAQQKNNL FGINAIDSSP GASANSFETV EGCINDFAKY
YISRGYSDPE DWRYFGGYLG NKGSGANVKY ASDPFWGEKA GQNAYIADYW TSGKGIAGLK
DYNYYQLGIY TGASSVTNKD NEKLYDVGSL YTERVEKIGA TTILTSKEKI NHNGKECYEI
NPVRTTPVIS NGSPVAFPGP YDWNDKGFVD ASKVKLINEG KYSENVMGKW ELNNGIWYYY
IDGKYVTGWN NIDNNKYYFD QSGKMQTGWL CIDNIWYHFR DNGTIDIGWR EIDGYWYYFN
NDGEMQKGWQ TIGKYKYYFN DNGVMNIGWR KIDNIWYYFN INGEMQTGWL FENNIWYHFE
TTGSMTVGWK KINNDWYYFN SNGEMQTGWQ NIGTPKYHFR DNGVMDIGWR EIDGQWHYFN
ENGEMQTGWQ YINGNSYYFD EKGIWKQ