Gene CPF_0616 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagCPF_0616 
Symbol 
ID4203949 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameClostridium perfringens ATCC 13124 
KingdomBacteria 
Replicon accessionNC_008261 
Strand
Start bp744000 
End bp745112 
Gene Length1113 bp 
Protein Length370 aa 
Translation table11 
GC content30% 
IMG OID638081501 
Productanaerobic sulfatase-maturase 
Protein accessionYP_695069 
Protein GI110800028 
COG category[R] General function prediction only 
COG ID[COG0641] Arylsulfatase regulator (Fe-S oxidoreductase) 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones13 
Plasmid unclonability p-value0.512735 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGCCACCAT TAAGTTTGCT TATTAAGCCA GCTTCTAGTG GATGTAATTT AAAATGCACT 
TATTGTTTTT ATCATTCTTT AAGTGATAAT AGAAATGTTA AGAGCTACGG AATTATGAGA
GATGAAGTTT TAGAAAGCAT GGTCAAAAGG GTTTTGAATG AAGCTAATGG ACATTGCAGT
TTTGCTTTTC AGGGAGGAGA ACCTACCTTA GCAGGATTAG AATTTTTTGA AAAGTTAATG
GAGCTTCAGA GAAAACATAA TTATAAAAAT TTAAAAATAT ATAATAGTTT GCAAACCAAT
GGAACTTTAA TAGATGAAAG TTGGGCAAAG TTTTTAAGTG AAAATAAATT TCTTGTGGGA
CTATCTATGG ATGGACCTAA GGAAATACAC AATTTAAATA GAAAAGATTG TTGTGGTTTA
GATACCTTTA GTAAGGTAGA AAGGGCAGCG GAGTTATTTA AAAAGTATAA GGTTGAATTT
AATATATTAT GCGTTGTGAC CTCTAATACA GCTAGGCATG TAAATAAAGT ATATAAATAT
TTTAAGGAAA AAGATTTTAA ATTTCTTCAA TTTATAAATT GTCTTGACCC ATTGTACGAG
GAAAAAGGTA AATATAATTA TTCTTTAAAG CCAAAGGATT ATACTAAGTT TTTAAAGAAT
TTATTCGACT TTTGGTATGA AGATTTTCTA AATGGAAATA GAGTAAGCAT TAGATATTTT
GATGGTTTAT TAGAAACAAT TTTATTAGGA AAGTCATCAT CTTGTGGAAT GAATGGGACA
TGTACCTGTC AGTTTGTTGT GGAAAGTGAT GGGAGTGTTT ATCCTTGTGA TTTTTATGTT
TTAGATAAAT GGAGACTAGG CAACATACAG GATATGACAA TGAAAGAATT ATTTGAAACC
AATAAAAATC ATGAGTTTAT AAAATTATCA TTTAAAGTTC ATGAAGAATG CAAAAAGTGC
AAGTGGTTTA GACTTTGTAA AGGTGGATGT AGAAGGTGCA GAGATTCAAA GGAAGATTCA
GCTTTAGAGT TAAACTACTA TTGTCAAAGC TACAAGGAAT TCTTTGAATA TGCCTTTCCA
AGGCTAATAA ATGTTGCCAA CAATATTAAA TAA
 
Protein sequence
MPPLSLLIKP ASSGCNLKCT YCFYHSLSDN RNVKSYGIMR DEVLESMVKR VLNEANGHCS 
FAFQGGEPTL AGLEFFEKLM ELQRKHNYKN LKIYNSLQTN GTLIDESWAK FLSENKFLVG
LSMDGPKEIH NLNRKDCCGL DTFSKVERAA ELFKKYKVEF NILCVVTSNT ARHVNKVYKY
FKEKDFKFLQ FINCLDPLYE EKGKYNYSLK PKDYTKFLKN LFDFWYEDFL NGNRVSIRYF
DGLLETILLG KSSSCGMNGT CTCQFVVESD GSVYPCDFYV LDKWRLGNIQ DMTMKELFET
NKNHEFIKLS FKVHEECKKC KWFRLCKGGC RRCRDSKEDS ALELNYYCQS YKEFFEYAFP
RLINVANNIK