Gene A2cp1_4070 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagA2cp1_4070 
Symbol 
ID7297495 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameAnaeromyxobacter dehalogenans 2CP-1 
KingdomBacteria 
Replicon accessionNC_011891 
Strand
Start bp4548182 
End bp4549147 
Gene Length966 bp 
Protein Length321 aa 
Translation table11 
GC content77% 
IMG OID643596877 
ProductDNA-formamidopyrimidine glycosylase 
Protein accessionYP_002494454 
Protein GI220919150 
COG category[L] Replication, recombination and repair 
COG ID[COG0266] Formamidopyrimidine-DNA glycosylase 
TIGRFAM ID[TIGR00577] formamidopyrimidine-DNA glycosylase (fpg) 


Plasmid Coverage information

Num covering plasmid clones21 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
GTGCCGGAGC TGCCGGACAT CGAGGTGTAC GTCGAGGCGC TCGCGGCGCG GGTGCTGGGT 
CAGCCGCTGG AGCGCATCCG GCTGGGGAAC CCCTTCCTGC TCCGCTCCGC AGACCCGCCG
CTCGCGGAGG CGGAGGGGCG GCGCGTCGCC GCGGTGCGCC GGCAGGGCAA GCGCCTCGTC
CTGGCGCTCG ACGGCGACCT CCACCTGGCG CTGCACCTCA TGATCGCGGG CCGGCTGCAC
TGGAAGGACC CCGGCGCGCG GCTCCCGGGG AAGGCGGGGC TGGCGGCGTT CGACTTCCCC
AACGGCACGC TGGTCCTCAC CGAGGCCGGC ACGAAGCGGC GCGCCGCGCT CCACCTGGTG
CGCGGCGCCG CGGCGCTCGC CGCGCTGGAC CGCGGCGGCA TCGAGCCGCT CGACGTGGAC
CTCGCCGCGT TCGCCGCCGC GCTCCGGCGC GAGAACCACA CGCTGAAGCG CGCGCTGACG
GATCCCTCGC TCTTCTCCGG CATCGGCAAC GCCTACTCGG ACGAGATCCT GCACCGGGCC
CGCCTGTCGC CGGTCGCGCT GACGTCGCGG CTCGGCGACG CAGAGGTGGC GCGCCTGTTC
GAGGCCACGC GCGAGGTGCT GACCGGCTGG ACGGCGCGGC TCCGCGAGGA GGCGGGGAGC
GGCTTCCCCG AGGGCGTCAC CGCGTTCCGC GAGGGAATGG CCGTGCACGG ACGGCACCGC
CAGCCGTGCC CGGTGTGCGG CACCGCGGTG CAGCGGATCG TGCGCGCGGA GAACGAGGTG
AACTACTGCC CGCGCTGCCA GACCGGCGGG CAGATCCTCT CGGACCGCTC CCTCGCCCGC
CTGCTGAAGC ACGACTGGCC ACGGACGGTG GACGAGCTGG AGCGCAACCC GGCGCTCGGC
CTCCGGCCCG CGCCGGGGCC GGCCGGGCCG CGATCCAGGG GGCCGCCGCG CCGTTCGTCG
CGCTGA
 
Protein sequence
MPELPDIEVY VEALAARVLG QPLERIRLGN PFLLRSADPP LAEAEGRRVA AVRRQGKRLV 
LALDGDLHLA LHLMIAGRLH WKDPGARLPG KAGLAAFDFP NGTLVLTEAG TKRRAALHLV
RGAAALAALD RGGIEPLDVD LAAFAAALRR ENHTLKRALT DPSLFSGIGN AYSDEILHRA
RLSPVALTSR LGDAEVARLF EATREVLTGW TARLREEAGS GFPEGVTAFR EGMAVHGRHR
QPCPVCGTAV QRIVRAENEV NYCPRCQTGG QILSDRSLAR LLKHDWPRTV DELERNPALG
LRPAPGPAGP RSRGPPRRSS R