Gene Franean1_5066 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagFranean1_5066 
Symbol 
ID5673402 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameFrankia sp. EAN1pec 
KingdomBacteria 
Replicon accessionNC_009921 
Strand
Start bp6065049 
End bp6066149 
Gene Length1101 bp 
Protein Length366 aa 
Translation table11 
GC content73% 
IMG OID641243917 
Productalcohol dehydrogenase 
Protein accessionYP_001509332 
Protein GI158316824 
COG category[C] Energy production and conversion 
COG ID[COG1062] Zn-dependent alcohol dehydrogenases, class III 
TIGRFAM ID[TIGR03451] mycothiol-dependent formaldehyde dehydrogenase 


Plasmid Coverage information

Num covering plasmid clones16 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones
Fosmid unclonability p-value0.0364575 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGACGGATG GCCCGCAGGC AGTGACGGTG GAGGCGGTCG TCGCGGTCGA GAAGGGCGCG 
CCGGTGGCGC TGACGCGGAT CATCGTGCCG CCGCCGGGCC CCGGGGAGGC CCGGGTCCGG
GTGCAGGCGT GCGGGGTGTG CCACACCGAC CTGCACTACC GCGAGGGCGC GATCAACGAC
GACTACCCGT TCCTACTCGG CCACGAGGCG GCCGGGACGG TCGAGTCCGT CGGCGATGGG
GTCACCTCGG TCGTCCCGGG TGACTACGTG GTGCTGGCCT GGCGGGCGCC GTGCGGGACG
TGCCGGTCGT GCCTGCGCGG GGCGCCCTGG TACTGCTTCG ACTCCCGCAA CGCCGTCAAC
CCGATGACGC TGCAGGACGG CACCCCGCTC TCCCCCGCTC TGGGCATCGG CGCCTTCACG
CCGCTGACAC TCGTGGCCGC CGGGCAGTGC GTCAAGGTCG ACCCGGCCGT TCCGCCGCAG
GCGGCGGGCC TGATCGGCTG TGGGGTCATG GCCGGCTTCG GCGCCGCGGT GAACACCGGA
CGGGTAACCC GCGGCGAGAC GGTCGCCGTC TTCGGCTGCG GCGGGGTCGG CGACGCGGCC
ATCGCCGGCG CGTCGGTCGC CGGCGCGCGC CGGATCATCG CCGTCGATGT GGACGACCGC
AAGCTCGAGT GGGCCCGCGG GTTCGGCGCG ACCCATGTCG TGAACTCCCG GAACGAGGAC
CCGGTGGAGG CCGTGCGGGC GCTGACCGAC GGCAACGGGG CGGACGTCGT GATCGAGGCG
GTCGGCCGCC CCGAGACCTA CCGGCAGGCA TTCTTCTCCC GCGACCTGGC CGGACGGCTG
GTGCTCGTCG GCGTGCCCGA CCCGTCGATG ACCGTCGAGC TGCCGCTCAT CGAGGTGTTC
AGCCGCGGCG GCTCGCTGGC GTCGTCCTGG TACGGCGACT GCCTCCCGAC CAGGGATTTC
CCGATCATCA TCGACCTGCA CCGCGGTGGC CGGCTCGACC TGGCCGCGTT CGTCACCGAG
ACGGTCGGCA TCGGGGACGT CGAGCGGGCC TTCGAGCGGA TGCGGCGCGG TGACGTGCTG
CGCAGCGTGG TCCTGATCTG A
 
Protein sequence
MTDGPQAVTV EAVVAVEKGA PVALTRIIVP PPGPGEARVR VQACGVCHTD LHYREGAIND 
DYPFLLGHEA AGTVESVGDG VTSVVPGDYV VLAWRAPCGT CRSCLRGAPW YCFDSRNAVN
PMTLQDGTPL SPALGIGAFT PLTLVAAGQC VKVDPAVPPQ AAGLIGCGVM AGFGAAVNTG
RVTRGETVAV FGCGGVGDAA IAGASVAGAR RIIAVDVDDR KLEWARGFGA THVVNSRNED
PVEAVRALTD GNGADVVIEA VGRPETYRQA FFSRDLAGRL VLVGVPDPSM TVELPLIEVF
SRGGSLASSW YGDCLPTRDF PIIIDLHRGG RLDLAAFVTE TVGIGDVERA FERMRRGDVL
RSVVLI