Gene Aazo_3417 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagAazo_3417 
Symbol 
ID9341222 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism name'Nostoc azollae' 0708 
KingdomBacteria 
Replicon accessionNC_014248 
Strand
Start bp3482403 
End bp3483554 
Gene Length1152 bp 
Protein Length383 aa 
Translation table11 
GC content36% 
IMG OID 
Productfamily 1 extracellular solute-binding protein 
Protein accessionYP_003722186 
Protein GI298492009 
COG category 
COG ID 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones14 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGAATCGAC GCTCTTTTTT GTTAGGTGCA AGCGGACTGG TATTTTCCCA GATACTCATG 
GGTTGTGCTG GTAAAAACCA GACACATCTA AATGTACAGT TATTAAAAGG TTCTATACCT
GGTCAGGTGG TTAATCAATT TCGTAAAACT CTAGAGTCAG ATGCAAATTT AAAGTTTGTC
CCTATTAATC AAATTTTAGA TTTATTTAAG CAATTACAAA TATGGCAAAA ACCAGAAACT
AAAGATCAAC AGGGATGGAA AAGTTATATT CCCTTGATGC AAAGTCAACA ACAATCTAAA
GCTGATTTAG TGACATTGGG AGATTTTTGG CTAAAAGCAG CAATTGAACA GAAACTGATT
CAACCACTAG AAACAGAAAA AATTAAACAA TGGTTAAGTT TAAATCTCAG GTGGCAGCAA
TTAGTAAGAC GCGATGATCA AGGAAATATA GATCCACAAG GAAAGATTTG GGCTGCACCT
TATCACTGGG GTAATACAGT CATTATTTAT AATGGGGAAA AATTTCAAAA ATTTGATTGG
CAACCAAAAG ACTGGAGCGA CTTATGGCGG AGTGAATTGC GATCGCGTAT TTCCTTACTT
AATCATCCCA GAGAAGTCAT TGGTTTGGTT TTAAAAAAAC TAGGAGAATC CTACAATACG
GAAAATATTA CTCAAATCCC CGACTTAAAA GCAGAATTAC TGGCACTACA CCAACAAGTA
AAATTTTATG ATTCTACTAC CTATCTAGAA CCACTACTCA CCGGAGATAC TTGGTTAGCT
GTGGGTTGGT CAAATGACGT TATGCCTATA CTCAGTCGTT ATCCAAAACT TACCGCAGTC
ATTCCCCAAT CAGGAACTGC AATGTGGGCA GACTTATGGG TAAGTCCGGC TGAAGTTGAG
CAAAACACCT TAGCATCTGA TTGGATTAAT TTTTGTTGGC AACCAAATAT AGCCAAACAA
ATTGCCATAC TGACTAAAAA TAATTCGCCT ATAACAAATA TTGTAGCCTC TGATCTTCAG
AAACCATTAC AAAACTTGTT ACTAAATAAT CAGGAATTAT TTGATAAAAG TGAATTTTTA
CTCCCCTTAC CAGCATCAGT CAATAAGGGT TATAAGTATT TATTTAACAA AATAAAAAAT
TCCGAACAAT GA
 
Protein sequence
MNRRSFLLGA SGLVFSQILM GCAGKNQTHL NVQLLKGSIP GQVVNQFRKT LESDANLKFV 
PINQILDLFK QLQIWQKPET KDQQGWKSYI PLMQSQQQSK ADLVTLGDFW LKAAIEQKLI
QPLETEKIKQ WLSLNLRWQQ LVRRDDQGNI DPQGKIWAAP YHWGNTVIIY NGEKFQKFDW
QPKDWSDLWR SELRSRISLL NHPREVIGLV LKKLGESYNT ENITQIPDLK AELLALHQQV
KFYDSTTYLE PLLTGDTWLA VGWSNDVMPI LSRYPKLTAV IPQSGTAMWA DLWVSPAEVE
QNTLASDWIN FCWQPNIAKQ IAILTKNNSP ITNIVASDLQ KPLQNLLLNN QELFDKSEFL
LPLPASVNKG YKYLFNKIKN SEQ