Gene Franean1_1992 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagFranean1_1992 
Symbol 
ID5670393 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameFrankia sp. EAN1pec 
KingdomBacteria 
Replicon accessionNC_009921 
Strand
Start bp2394912 
End bp2396000 
Gene Length1089 bp 
Protein Length362 aa 
Translation table11 
GC content73% 
IMG OID641240913 
Productshort-chain dehydrogenase/reductase SDR 
Protein accessionYP_001506335 
Protein GI158313827 
COG category[I] Lipid transport and metabolism
[Q] Secondary metabolites biosynthesis, transport and catabolism
[R] General function prediction only 
COG ID[COG1028] Dehydrogenases with different specificities (related to short-chain alcohol dehydrogenases) 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.00359465 
Plasmid hitchhikingNo 
Plasmid clonabilitydecreased coverage 
 

Fosmid Coverage information

Num covering fosmid clones
Fosmid unclonability p-value0.191358 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGTCCGACT TCATATGTCG CGACTGGATT TTGATCCTTA CGTCGGCACG TTCATCGCCC 
ATACCCCGCT GCGGATGTTC GTGCTGGGTC CGGAGAAGGG CGCGCTCGGT CTCGGTCAAC
GGCGCGGCGA TCTCCTGCGA CGGCACGGAC GTCGACGGCG CCTATCCCGT CGTGCTTGTC
CGGCCCACCG CACAGGGCTT CGGGTCCGCG GAGGTCGCGG GTCCGAATAG GGTTGCGGGC
AGTGGCGGAG GGGGCGCGCT GGGCTGGACT GGGCGGACCG TGGGCGGCAG AGGAGACGGC
GTGGATCTCG AGCTGGCAGG CAAGGTCGTC CTGATCACCG GCGGCTCGGA CGGGCTCGGC
GCCGCGTTGG TGCGTACCCT CGCCGCGGAG GGGGCGCGTG TCGCGTTCTG CGCCCGCGAC
GCGGACCGCC TCAACCGCCT CGCTGCCGAG GTGGCACCGA CCGCCGCGGC CGGCGCGGAG
TTCCTGCCCG TCCCCGCGGA CGTCACCCGC CTGGCCGACC TCGAGCGGTT CGTCGAGCAG
GCCGTCGGCC GCTGGGGGCG GATCGACGGC CTGGTGAACA ACGCCGGGCG CAGCGCCGCC
GGGCCGTTCG CGTCGCACAC CGACGAGGTC TGGGACGCCG ACCTGCAGCT CAAGGTGCAC
AGCACCGTCC GGCTGACCCG GTTGGCGCTG CCGCACCTGC GCGCGGCCGG CGGCGGCTCC
GTGATCAACA CGCTGGCCAT TGCCGCGAAG ACCCCCGGCG CCGGATCGAC CCCCACCTCG
GTCTCCCGGG CGGCCGGGCT CGCCCTCACC AAGGCGCTGT CCAAGGAACT CGGCCCGGAC
GGAATCCGGG TGAACGCGGT GCTGATCGGG CTGCTGGAGA GCGGCCAGTG GGACCGCCGC
GCAGCCGAGC AAGGCATCGG CGTGGACGAG CTGTACGCCG AGCTGAGCCG AGGCTCCGAC
ATCCCCCTCG GACGGGTCGG CCGCGCCCAG GACTTCGCCG ACCTCGCGGC CTTCCTGCTC
TCCCCCCGCG CCGGTTACCT GACCGGCGTC GGCATCAACC TCGACGGGGG CCTCTCGCCG
GCGCCCTGA
 
Protein sequence
MSDFICRDWI LILTSARSSP IPRCGCSCWV RRRARSVSVN GAAISCDGTD VDGAYPVVLV 
RPTAQGFGSA EVAGPNRVAG SGGGGALGWT GRTVGGRGDG VDLELAGKVV LITGGSDGLG
AALVRTLAAE GARVAFCARD ADRLNRLAAE VAPTAAAGAE FLPVPADVTR LADLERFVEQ
AVGRWGRIDG LVNNAGRSAA GPFASHTDEV WDADLQLKVH STVRLTRLAL PHLRAAGGGS
VINTLAIAAK TPGAGSTPTS VSRAAGLALT KALSKELGPD GIRVNAVLIG LLESGQWDRR
AAEQGIGVDE LYAELSRGSD IPLGRVGRAQ DFADLAAFLL SPRAGYLTGV GINLDGGLSP
AP