Gene Franean1_2743 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagFranean1_2743 
Symbol 
ID5671134 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameFrankia sp. EAN1pec 
KingdomBacteria 
Replicon accessionNC_009921 
Strand
Start bp3245050 
End bp3246375 
Gene Length1326 bp 
Protein Length441 aa 
Translation table11 
GC content75% 
IMG OID641241655 
Producthypothetical protein 
Protein accessionYP_001507075 
Protein GI158314567 
COG category 
COG ID 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.0118052 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones15 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGACACATC CGGTAAGGAT TCCCGCTGAT GTTGATCGTG AGGACAGGAT CGTGGCGAAT 
TTGACCGCCC GCCAGGTCCT GATCCTCACG CTCACCGGCA CCGTGCTCTA CCTGGGCTGG
GCCGCGACCC GCGCCCTGCT GCCGCTGCCG CTGCCGGTGT TCGCCGTGCT GGCGGTGCCG
GTCGCCGTGG GCGCCGGTGT CCTCGTCCTC GGCCAGCATG ACGGGCTGTC CCTGGACCGG
CTGCTGGTCG CCGCGATCCG CCAGCGCACC AGCCCGCGGC ATCGGATCAA CGCCCCCGAA
GGGGTGATCG CGCCGCCGTC GTGGCTGGCC GCCCGCGCCA CCAGCGGCCC CGATGAACGG
AGACCGGCCG CCAGCGGTCA GAGCGCGGTG CCGCTACGGC TGCCGGCCCG CACCGTCACC
GGCCAGGCCG GGGTCGGGGT GATCGACCTC GGCGGGGACG GCCTGGCGGT GGTCGCGGTC
GCGAGCACGG TGAACTTCGC GTTGCGCACG CCGGGGGAGC AGGACGGGCT GGTCGCCGTG
TTCGCCCGCT ACCTGCACTC CCTCACCGCG CCGGTGCAGA TCCTGGTGCG GGCGATGCCC
GCGGACCTGT CCGGTCAGAT CCACCTGCTC GACGACGCCG CCGAGCAGCT GCCCCACCCG
GCGCTCGCGC ACGCCGCCCG CGAACACGCC ACCTACCTGG GCCAGCTCGC TGTCGAGATG
CAGCTGCTGA CCCGTCAGGT GCTGCTGGTG CTGCGTGAAC CACTCGTCAC CGCCGGCCCG
GTCGATGGGC TCGGGGGCGC GTCCCCGCTG GCCGCGTGGA CGGGCCGGCG GGCGGCGGTC
CGCGACGCCC GCCGTGCCGG AGCCGCTGCC CGTCGTGCCG CGCACACCCG GCTCACCCGC
CGCCTCGCCG AGGCCACCGA CCTGCTCGCC CCCGCCGGGA TCGTCATCAC CCCCCTGGAC
GCCGGCACGG CGACCAGCGT GCTGGCCGCC GCCTGCAACC CCGCCGGCCT GGTACCACCG
GCCGCGCTCG CCGCGCCCGA CGAGGTCATC ACCGCCGACT TTCCCGAGCC CACCGACAGC
TACCCGGCCT ATCCGCCGGA CACCGACGAC GGCGGCTTTC CGGACGACGC CGGGTTCGAC
GACCCGGGCG CGGCTGTCGG CCCCGGCTAC GACGACCGGT TCGACGACGC GGACGGGGAC
GGCCTGTCCG ACGCCGACGA CCCGGACTTC TGGGACCCAC CCGTCCGCCG CCCGCCGGCC
GGGCGATCCG AGGGCGGCTC CCGACGGCCA GCACGACACA CGCGACGCAG GGGGCAGGCC
CGATGA
 
Protein sequence
MTHPVRIPAD VDREDRIVAN LTARQVLILT LTGTVLYLGW AATRALLPLP LPVFAVLAVP 
VAVGAGVLVL GQHDGLSLDR LLVAAIRQRT SPRHRINAPE GVIAPPSWLA ARATSGPDER
RPAASGQSAV PLRLPARTVT GQAGVGVIDL GGDGLAVVAV ASTVNFALRT PGEQDGLVAV
FARYLHSLTA PVQILVRAMP ADLSGQIHLL DDAAEQLPHP ALAHAAREHA TYLGQLAVEM
QLLTRQVLLV LREPLVTAGP VDGLGGASPL AAWTGRRAAV RDARRAGAAA RRAAHTRLTR
RLAEATDLLA PAGIVITPLD AGTATSVLAA ACNPAGLVPP AALAAPDEVI TADFPEPTDS
YPAYPPDTDD GGFPDDAGFD DPGAAVGPGY DDRFDDADGD GLSDADDPDF WDPPVRRPPA
GRSEGGSRRP ARHTRRRGQA R