Gene Franean1_4681 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagFranean1_4681 
Symbol 
ID5673023 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameFrankia sp. EAN1pec 
KingdomBacteria 
Replicon accessionNC_009921 
Strand
Start bp5590958 
End bp5592436 
Gene Length1479 bp 
Protein Length492 aa 
Translation table11 
GC content71% 
IMG OID641243538 
Productglycoside hydrolase family protein 
Protein accessionYP_001508954 
Protein GI158316446 
COG category[G] Carbohydrate transport and metabolism 
COG ID[COG3507] Beta-xylosidase 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.026787 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones15 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGAACGCCA ACCCGATCCT GCCCGGTTTT CACCCGGACC CGTCGATCTG CCGGGTGGGC 
GACGATTACT ACCTGGTGAC CTCGAGCTTC GAGTACTTCC CCGGGGTGCC GATCTTCCGC
AGCACGGACC TCACCGGCTG GGAACAGATC GGCAACGTGC TCGACCGCCC CACCCAGCTG
AACGTGACCC CGGGCCTGGA GTCGGCCAGC ACAGGGATCT TCGCGCCGAC CCTGCGCCAC
CATGACGGCA GATTCTGGCT CGCCACGACG AACTTCACCG ACGTCCGTAA GGGTCACCTC
ATCGTCCAGG CGGCCGACCC GGCCGGCGCA TGGACCGACC CGGTCCACAC AGGCGGGGGA
ACGGTCGGGA TTGACCCCGA CCTGGCGTGG GACGAGAACG GTACCTGTCA TCTGACGTGG
TGCTTTCCCG GGCAGATCAT GCAGGCCGCC GTCGACCCGG AGAGCGGGAG GCTGCTCTCC
GAGCCCCGTG GGCTGTGGAG CGGGACGGGG CGTGCCCACC CTGAAGGCCC GCACCTGTTC
AGCCGGGAAG GCTGGTGGTA CCTGGTCATC GCCGAGGGCG GCACCGACGG TGGCCACGCC
GTGTCGATCG CGCGTTCCCG CTCGATCACC GGACCCTTCA CCGGCAACCC CGCCAATCCG
ATCCTCACCA GGAGCGGCAC CGAACACCCG GTCCAGAGCA CCGGCCACGC CGACTTCGTC
GAACTTCCCG GGGGTGAGTG GGCGATGGTT CACCTGGGGG TCCGCCCGCG CGGCACGTTC
CCCAAGTTCC ACGTCAACGG CCGGGAGACC TTCCTGACCG GCATTACCTG GGCCGACGGC
TGGCCGATCG TGGTGGAGGA CCGGTTCACG GTGCCAGTCC GGGACAACTC CTTCGTCGAC
GAGTTCCGCA CGCCGACACT GCATCCCCGG TGGGTCTCCC CGGGGACCGA CCCACGGACT
TTCACCCGCC ACCGACCCGG CGGAGTCGTC CTGGCCGCCG GCCGGGCACC CGACGCGGGC
GAGGCCAGGC GCCTGCTCGC TGTGCGCGCC CAGGACCCGC AATGGCAGGT AACCGCGGTC
ATCCCCGACG GCGACGCCTG CCTGACCGTC CGGATGGACG ACGCCCACTG GGTCGCCGTC
GAACGCCGTG GGGAGATGCT GGCGGCCAGG ATGGTGCTCG GCCCACTTGA CCAGACCCTC
GCCACCGCCG CCGGGATCGG CCCGGGCGAC GCGCTCGCCG TCCGCGCCGT GACGCATGCC
GAGGCCGCCA GCTTCCGTGC CGGCCCCGAC CAGCTCGAGC TGGGCCACCT CACCGACGGC
GAGTTCCGGC TGCTCGCCAC GGTCGACGGG CGCTACCTCT CGACCGAGGT CGCCGGCGGC
TTCACCGGGC GCGTCGTCGG AGTCGAGGCG ATCGGCACCG ACGCCACCCT GTCACGATTC
GAATACCTGG CCTTGGATGC CGCCGCGCCC CAGGGCTAG
 
Protein sequence
MNANPILPGF HPDPSICRVG DDYYLVTSSF EYFPGVPIFR STDLTGWEQI GNVLDRPTQL 
NVTPGLESAS TGIFAPTLRH HDGRFWLATT NFTDVRKGHL IVQAADPAGA WTDPVHTGGG
TVGIDPDLAW DENGTCHLTW CFPGQIMQAA VDPESGRLLS EPRGLWSGTG RAHPEGPHLF
SREGWWYLVI AEGGTDGGHA VSIARSRSIT GPFTGNPANP ILTRSGTEHP VQSTGHADFV
ELPGGEWAMV HLGVRPRGTF PKFHVNGRET FLTGITWADG WPIVVEDRFT VPVRDNSFVD
EFRTPTLHPR WVSPGTDPRT FTRHRPGGVV LAAGRAPDAG EARRLLAVRA QDPQWQVTAV
IPDGDACLTV RMDDAHWVAV ERRGEMLAAR MVLGPLDQTL ATAAGIGPGD ALAVRAVTHA
EAASFRAGPD QLELGHLTDG EFRLLATVDG RYLSTEVAGG FTGRVVGVEA IGTDATLSRF
EYLALDAAAP QG