Gene Aazo_2041 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagAazo_2041 
Symbol 
ID9339834 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism name'Nostoc azollae' 0708 
KingdomBacteria 
Replicon accessionNC_014248 
Strand
Start bp2119848 
End bp2121437 
Gene Length1590 bp 
Protein Length529 aa 
Translation table11 
GC content43% 
IMG OID 
Productradical SAM domain-containing protein 
Protein accessionYP_003721221 
Protein GI298491044 
COG category 
COG ID 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones13 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGAATGTAT TACTTATATA TCCGTTGTTT CCAAAAAGTT TTTGGTCTTT TGAAAAAACA 
CTAGCTTTGC TAGACAGGAA AGCGATGTTA CCACCATTGG GCTTGGTGAC AGTAGCCGCA
ATTTTACCCC AACAATGGAA TTTTAAGCTA GTAGACAGGA ATATTCGCCA AATTACCGAA
GCAGAATGGG CTTGGGCTGA TTTGGTGATT TTATCAGCGA TGATTGTCCA AAAAGAGGAT
TTACTCGCAC AGATTCAGGA AGCAAAGCGT CGTGGTAAGC TTGTGGCTGT GGGTGGACCA
TACCCGACAG CATTACCTAA CGAAGTCACA GATGTGGGAG CAGATTATTT GATTTTGGAT
GAAGGGGAAA TTACGCTACC TTTATTTATA GATGCGATCG GACGCGGTGA ATCTTCAGGA
ATCTTTCGTT CTGGTGGTGA AAAACCAGAT GTGACAAACA CTCCCATTCC TCGTTTTGAC
CTACTGGAAT TTGATGCCTA TGCGGAAATG TCAGTGCAAT TTTCCCGTGG CTGTCCCTTC
CAGTGTGAGT TCTGCGACAT TATCGTCCTC TACGGTCGCA AACCCCGCAC CAAAACACCA
GCCCAACTCC TCGCAGAACT TGATTATCTC TATGAATTAG GTTGGCGACG CAGTATTTTT
ATGGTGGATG ATAACTTCAT CGGCAATAAG CGTAACGTTA AATTATTCTT GAAAGAACTA
CAACCTTGGA TGGTTGCACA TCATTATCCC TTCTCCTTTG CCACAGAGGC TTCCGTTGAC
TTAGCCCAAG ATCAAGAATT GATGGATGCA ATGGTAAGGT GTAATTTCGG GGCTGTGTTC
TTGGGAATTG AAACCCCCGA CGAAGAAAGC CTTACTTTTA CTCAAAAATT CCAAAATACT
CGAGATTCCC TCACCGAAGG AGTAAATAAA ATTACACGCT CAGGGTTACG AGTCATGGCA
GGTTTTATTA TTGGCTTTGA TGGCGAAAAG TCGGGTGCTG GGGCGAGAAT TGTTAAATTT
GTGGAACAGA CAGCCATTCC CACCGCTTTA TTTAGTATGC TCCAGGCCCT GCCTGATACA
GCATTGTGGC ATAGATTAGA AAAGGAAAAC CGACTCCGCA ATAAATCTGC TAATATTAAT
CAAACCACAT TGATGAATTT TTTTCCTACA AGACCTTTAG CAGAAATAGC CAGCGAATAT
GTAGAAGCGT TTTGGGAACT ATACGAACCA TCAAGATTTT TAGATCGTGC TTATCGGCAC
TACCGCATTT TGGGTCAAGC AACCTATCCC AAAAAGGGCA AAGGTGCTAA AAAACCATTG
AATTGGAAGG TACTGCGGGC ATTGTTGACT ATTTGCTGGC AACAAGGTGT GTTCCGTAAT
ACTCGTTGGC AATTCTGGCG CAATCTCTGG AGTATGTACA AGCATAATCC TGGTGGTATC
AGTAGTTATT TAGCCGTTTG CGCTCAAATT GAGCATTTCT TGGAATATCG TCAGATTGTG
CGGGATCAAA TTGAGGTTCA AATGGCTGAG TTTTTGGCAG CAGAAGCTCA AGTTAAGCTT
GAGGAAGAAA AAGCTCAAGT TTTAGTTTAG
 
Protein sequence
MNVLLIYPLF PKSFWSFEKT LALLDRKAML PPLGLVTVAA ILPQQWNFKL VDRNIRQITE 
AEWAWADLVI LSAMIVQKED LLAQIQEAKR RGKLVAVGGP YPTALPNEVT DVGADYLILD
EGEITLPLFI DAIGRGESSG IFRSGGEKPD VTNTPIPRFD LLEFDAYAEM SVQFSRGCPF
QCEFCDIIVL YGRKPRTKTP AQLLAELDYL YELGWRRSIF MVDDNFIGNK RNVKLFLKEL
QPWMVAHHYP FSFATEASVD LAQDQELMDA MVRCNFGAVF LGIETPDEES LTFTQKFQNT
RDSLTEGVNK ITRSGLRVMA GFIIGFDGEK SGAGARIVKF VEQTAIPTAL FSMLQALPDT
ALWHRLEKEN RLRNKSANIN QTTLMNFFPT RPLAEIASEY VEAFWELYEP SRFLDRAYRH
YRILGQATYP KKGKGAKKPL NWKVLRALLT ICWQQGVFRN TRWQFWRNLW SMYKHNPGGI
SSYLAVCAQI EHFLEYRQIV RDQIEVQMAE FLAAEAQVKL EEEKAQVLV