Gene Aazo_1032 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagAazo_1032 
Symbol 
ID9338827 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism name'Nostoc azollae' 0708 
KingdomBacteria 
Replicon accessionNC_014248 
Strand
Start bp1103808 
End bp1105088 
Gene Length1281 bp 
Protein Length426 aa 
Translation table11 
GC content38% 
IMG OID 
Productpeptidase M16 domain-containing protein 
Protein accessionYP_003720518 
Protein GI298490341 
COG category 
COG ID 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones12 
Plasmid unclonability p-value0.856018 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGACATCAA CTCTGTTAAA ATTTCCTCGA CTTAATGCTC CCAAGGTGCA TCATTTACCA 
AATGGTTTAA CCATCATCGC CGAGCAAATG CCAGTACCAG CAGTTAACCT TAATCTATGG
GTAAACATCG GTTCTGCTGT GGAGTTAGAT GCTATTAATG GCATGGCTCA TTTTTTAGAA
CATATTGTTT TTAAGGGAAC AGAGAGACTA GCCAGTGGTG AGTTTGAACG TCGAATTGAA
GAACGCGGCG CTGTTACTAA TGCCGCTACT AGCCAAGATT ATACACACTA TTATATTACG
ACTGCGCCGA AAGACTTTGC AGAATTAGCA CCATTGCAAA TAGATGTTGT TTGTAATCCT
AGTATTCCTG ATGATGCTTT TGAGAGAGAA CGCTTAGTAG TTTTGGAAGA AATCAGACGT
TCACAAGATA ACCCCAGACG GCGGATTTAT CGCCGCACAA TGGAAACCGC TTTTGATGTT
TTACCTTATC GTCGTCCGGT ACTCGGTCCA GAAGCAGTAA TTTCTCAAGT TACACCTCAG
CAAATGCGAG ATTTTCACCA TACCTGGTAT CAACCGTCTT CTATAACTGC GGTTGCCGTC
GGTAATCTAC CAGTAGAAGA ATTAATAGAA ATTATTGCCG AAGAATTTAG TAAAAATAGT
CAAAAATCAA AAATTAATAA TCAACAATTA ACCGTTAGTC AAGAACCTGC ATTTACAGAA
ATTGTGCGTC GGGAATTTAC TGATGAGAGT GTACAACAAG CCAGATTAAT AATCCTATGG
CGAGTTCCTG GACTCATGGA ATTAGATGAA ACATACTCTT TAGATGTGTT AGCAGGAATT
TTAGGACATG GACGTACATC TAGATTAGTC CATGATTTGC GAGAAGAAAG AGGACTTGTT
TCCTCAATTG CTGTTAGTAA TATTAATAAT CGACTGCAAG GGATATTTTC TATTTCTGCT
AAGTGTGAAG TAGATGATTT AGAAGCAGTA GAAGCTGCAA TTGCTAAACA TTTGTATACA
ATACAAACAG AATTAGTAAA AGAATCAGAA ATTTATCGTG TACGGCGACG GGTAGCCAAT
CGGTTTATAT TTGGGAATGA AACACCAAGT GAGCGCTCCG GTTTGTATGG TTATTATCAA
TCTTTAATAG GCGACCTAGA AGCAGCATTT AATTATCCCC AATATATACA AGCTCAAAAT
ACAAATAACT TAATCCAAGC TGCACAGAAA TATCTTGACC CCAACGCTTA TGGTGTAGTT
GTGATCAAAC CTGTTAAGTG A
 
Protein sequence
MTSTLLKFPR LNAPKVHHLP NGLTIIAEQM PVPAVNLNLW VNIGSAVELD AINGMAHFLE 
HIVFKGTERL ASGEFERRIE ERGAVTNAAT SQDYTHYYIT TAPKDFAELA PLQIDVVCNP
SIPDDAFERE RLVVLEEIRR SQDNPRRRIY RRTMETAFDV LPYRRPVLGP EAVISQVTPQ
QMRDFHHTWY QPSSITAVAV GNLPVEELIE IIAEEFSKNS QKSKINNQQL TVSQEPAFTE
IVRREFTDES VQQARLIILW RVPGLMELDE TYSLDVLAGI LGHGRTSRLV HDLREERGLV
SSIAVSNINN RLQGIFSISA KCEVDDLEAV EAAIAKHLYT IQTELVKESE IYRVRRRVAN
RFIFGNETPS ERSGLYGYYQ SLIGDLEAAF NYPQYIQAQN TNNLIQAAQK YLDPNAYGVV
VIKPVK