Gene Francci3_3497 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagFrancci3_3497 
Symbol 
ID3905231 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameFrankia sp. CcI3 
KingdomBacteria 
Replicon accessionNC_007777 
Strand
Start bp4170290 
End bp4171909 
Gene Length1620 bp 
Protein Length539 aa 
Translation table11 
GC content68% 
IMG OID637880819 
ProductXaa-Pro aminopeptidase 
Protein accessionYP_482579 
Protein GI86742179 
COG category[E] Amino acid transport and metabolism 
COG ID[COG0006] Xaa-Pro aminopeptidase 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones14 
Plasmid unclonability p-value0.434008 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones12 
Fosmid unclonability p-value0.658435 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGAGTGCCC AGTCCCACCA AAGCTCACCT CCACCGACCG GCACGGGGGT GCCCGAGAGC 
GCGCCGGCCG GGCGTGATCC GAACCTCATC ACGACCGACG AGCCGGGCGT TGTCCCCACC
GGGGCCAGTG AACCTCCGGC ATCGGATGAT TCCGTGCGCG AACCCGGACA GGCCAACCAC
GACGAGGAAC CACACGCGGC GTTCCGGCGA TTTATGTCGA CCGGGTGGGC GCCGATTGAC
GACGCTGTGC CATCCCGCGA TGACTGCGCA CCCTACACCA CGAAACGTCG CGCCTTGTTG
GGTTCGCGGT TCCCGGCGGA CGCGCTTGTG GTGCCGAGCG GCGGTCTGCG GGTACGGGCA
AACGACACAG ATTTTCCGTT CCGTCCCGGA AGTGACTTCT TCTGGCTGAC CGGCTGCCAT
GAACCCGATG CCGTGTTGGT TCTGCTACCC AACGCGGCCG GTGAGCATGA TGCGGTGCTC
TATGTCGCCG GGCGGTCCGA CCGGTCGAGC TCCGCCTTCT ACACCGACCG GCGATACGGG
GAACTGTGGG TCGGTCCCAG ACCGGGGGTT CGGGAGACCA CGGCACAGCT GGAGATCGAA
TGCCGGCCGC TCACGGAGCT CGCGGACGCG CTCGCGCGGT TGGCACCGGG AAACACCCGG
GTACTACGCG GAGTCGATCC GGTCGTGGAC GGTTCGGTGC TGCGCTGGTC CGGCCGGGGT
GGACCAGGAA CCGACCGCGA TCTGGCCCTC GCACAGGCCC TTTCCGAGCT GCGGCTGCTC
AAGGACGACT TCGAGATCGC CCGATTGGAC GAGGCGATCG CCGCGACCGT GCGTGGCTTC
ACCGAGTGCG TGGGAGAAAT CGGTCGGGCG ATGGACCTGC CCAACGGAGA ACGATGGTTG
GAGGGGACGT TCTGGCGGCG GGCCCGGGTG GACGGCAACG ACGTGGGATA CGGTTCCATC
GTCGCCTGTG GCCACCACGC CACGACGCTG CACTGGGTCC GCGACGACGG GCCGGTACGC
GCCGGTGATC TCGCCCTGCT CGACATGGGG GTCGAGGGAC GGTCCCTCTA CACGGCCGAC
GTCACCCGCA CGCTTCCGAT CAACGGACGG TTCACCGAGG TGCAACGCCA GGTCTATGAC
GTGGTTCTTC GCGCCCAGCA GGCCGGGATC GACGCGGTCC GGCCGGGAGC GTCCTTCCTG
GAGCCGCACC GGGCCGCCAT GCGGGTCATC GCCACGGCGC TCGACGACTG GGGGCTCCTG
CCGGTCAGTG CGCAGGAGTC GCTGACCGAG GACCCACGCT CCCCGGGAGC GGGCGTCCAT
CGCCGCTATA CCCTGCACTC CACGTCGCAC ATGCTCGGAT TGGACGTGCA CGACTGCGCG
CAGGCGCGCG ACCAGACCTA TCGCGACGCG GCGCTGGAAC CGGGCATGGT ATTGACGGTC
GAACCGGGCC TGTACTTCCA GCCGGACGAC CTGATGGTTC CGCCGGAACT GCGAGGCATC
GGCGTCCGGA TAGAAGACGA CATCCTGGTC ACCAGGGATG GAAGGCGGAA CATGTCCGCC
GCACTTCCGC GAACCGCGGA AGACGTCGAA AACTGGATGG CCCGGCAGCT GCCGAACTGA
 
Protein sequence
MSAQSHQSSP PPTGTGVPES APAGRDPNLI TTDEPGVVPT GASEPPASDD SVREPGQANH 
DEEPHAAFRR FMSTGWAPID DAVPSRDDCA PYTTKRRALL GSRFPADALV VPSGGLRVRA
NDTDFPFRPG SDFFWLTGCH EPDAVLVLLP NAAGEHDAVL YVAGRSDRSS SAFYTDRRYG
ELWVGPRPGV RETTAQLEIE CRPLTELADA LARLAPGNTR VLRGVDPVVD GSVLRWSGRG
GPGTDRDLAL AQALSELRLL KDDFEIARLD EAIAATVRGF TECVGEIGRA MDLPNGERWL
EGTFWRRARV DGNDVGYGSI VACGHHATTL HWVRDDGPVR AGDLALLDMG VEGRSLYTAD
VTRTLPINGR FTEVQRQVYD VVLRAQQAGI DAVRPGASFL EPHRAAMRVI ATALDDWGLL
PVSAQESLTE DPRSPGAGVH RRYTLHSTSH MLGLDVHDCA QARDQTYRDA ALEPGMVLTV
EPGLYFQPDD LMVPPELRGI GVRIEDDILV TRDGRRNMSA ALPRTAEDVE NWMARQLPN