Gene Franean1_7046 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagFranean1_7046 
Symbol 
ID5675357 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameFrankia sp. EAN1pec 
KingdomBacteria 
Replicon accessionNC_009921 
Strand
Start bp8597399 
End bp8598604 
Gene Length1206 bp 
Protein Length401 aa 
Translation table11 
GC content61% 
IMG OID641245892 
Productinulin fructotransferase (DFA-I-forming) 
Protein accessionYP_001511283 
Protein GI158318775 
COG category 
COG ID 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones16 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones19 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
TTGTCTACCC TCTATGACGT GACCACCTGG ACCGTCCCCA GCAACCCCTC CATCACCTGC 
TACGTCGACG TCGGCGTGGT AATCAACGAC ATCATCCGCG ACATCAAGGC CCAGCAGCCC
AACCAGTCGG CGAAGCCCGG AGCCGTCATC TACATCCCAC CGGGGGACTA CCCCCTGAAG
ACACGGGTGA CCGTCGACAT CAGCTACCTG ACCATCAAGG GATCCGGCCA CGGCTTCACC
TCGTCCAGCA TACGGTACAA CACGTCCAAC ACCTCGGCGT GGCACGAGTT ATGGCCGGGA
AGCAGCCGTA TCAAGGTCGA GAACACCGAC GGCAACAGCG AGGCGTTTCT GGTGTCCCGG
ACGGGAGACC CACGACTGAG CTCGGTCGTG TTCTCGAACT TCTGCCTCGA TGGCCTCAGC
TTCGGCACCA ACCAGAACTC GTATGTCAAC GGAAAGACCG GAGTTCGGGT CGCCACAAGC
AACGACGCTT TCAGGTTCGA GGGGATGGGC TTCGTCTACC TCGAGCACGC ACTGATAGTC
ACGAACGCCG ACGCACTGAG CGTCAGCGAC AACTTCATCG CCGAATGCGG GAGCTGCATC
GAGCTGACCG GCTCGGGACA GGCGTCCAAA ATCACCGACA ACCTGATCGG TGCGGGTTAC
GTCGGCTACT CGGTGTTCGC CGAGGGACAC GAGGGTCTGC TTGTCTCCGG TAACAACATC
TTCCCCCGGG GAAGGAGCGC GGTGCACTTC AAGAACACCA ACCGCTCAAC GATCACAGCC
AACCGTCTGC ACGACTTCTA CCCGGGGATC ATCGACTTCG AGGGGCTGAA CAAGGAGAAC
CTGATCAGCA GCAACCACTT CCGGCGCGAG GCTGAGCCGT GGCCTCCGAT GCAGTCCTAC
AACAACGGCA AGGACGACCT GTACGGGCTG GTGCACCTGC GCGGGGACAA CAACATGGTC
TCCACGAACC TCTTTGCCTT CTACGTCGAC CCCAGCAAGA TCACCCCCCT CGGCGCCACA
CCAACGATCA TCCTGGTCGC GTCCGGAAAC GGGAACTTCA TCTCCAACAA CCACGTCACG
GCCAACGTGG GCGTAAAGAA CGTCGTCCTC GACGCGACGA CGACCGGCAC GAAGGTTCTG
GACAGCGGGA CGGCGTCCGA GTTCGTTTCG TACACCTCCA ACTACACCTT CCGCCCGACG
CCGTGA
 
Protein sequence
MSTLYDVTTW TVPSNPSITC YVDVGVVIND IIRDIKAQQP NQSAKPGAVI YIPPGDYPLK 
TRVTVDISYL TIKGSGHGFT SSSIRYNTSN TSAWHELWPG SSRIKVENTD GNSEAFLVSR
TGDPRLSSVV FSNFCLDGLS FGTNQNSYVN GKTGVRVATS NDAFRFEGMG FVYLEHALIV
TNADALSVSD NFIAECGSCI ELTGSGQASK ITDNLIGAGY VGYSVFAEGH EGLLVSGNNI
FPRGRSAVHF KNTNRSTITA NRLHDFYPGI IDFEGLNKEN LISSNHFRRE AEPWPPMQSY
NNGKDDLYGL VHLRGDNNMV STNLFAFYVD PSKITPLGAT PTIILVASGN GNFISNNHVT
ANVGVKNVVL DATTTGTKVL DSGTASEFVS YTSNYTFRPT P