Gene Aazo_2958 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagAazo_2958 
Symbol 
ID9340762 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism name'Nostoc azollae' 0708 
KingdomBacteria 
Replicon accessionNC_014248 
Strand
Start bp3041032 
End bp3042360 
Gene Length1329 bp 
Protein Length442 aa 
Translation table11 
GC content40% 
IMG OID 
Productpheophorbide a oxygenase 
Protein accessionYP_003721890 
Protein GI298491713 
COG category 
COG ID 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.0467803 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGCAAGCTG AATTCAACTT TTTCCACCAC TGGTATCCTC TCTCACCAGT TGAGGACCTT 
GATTCACCGC GACCTGTTCC CGTAACTTTA TTAGGAATTC GATTAGTTAT ATGGAAGCCT
AGAGATGCAG ACAATTACCG TGTATTTTTA GATCAGTGTC CTCACCGTCT TGCACCCTTG
AGTGAAGGAA GAGTAGACGA TAAAACTGGG AATCTCATGT GTAGTTATCA CGGTTGGCAG
TTTGATGGTC GGGGTATTTG TACTCACATT CCTCAAGCTG AGAATCCTCA ACTTGTAGCT
AAAAATCAAC AAAATTTCTG TGTACTTTCC CTACCAGTGC GGCAAGAAAA TGATTTACTC
TGGGTTTGGC CTGATGCTAA ATCAGCGGAA CTAGCTGCAA CTACACCCCT ACCTTTATCA
CCACACATAG ATACTAACAA AGGTTTTGTC TGGTCTTCTT ATATTCGTGA CTTAGAATAT
GATTGGCAAA CCTTAGTAGA AAATGTAGCA GATCCTAGTC ATGTTCCCTT TTCTCATCAT
GGGGTACAGG GTAATCGTGA CAAAGCAACA TCCATTCCTC TTAATGTTGT CCAATCAACA
ATCAATTTAA TTGAAGTTTC CATTTCCAAA GCCTTGCCCA CAACAATCAC TTTTCAACCA
CCTTGTCTGT TAGAGTATGC AATTAGTATT GGTGACACTG ACAAGAAATT AGGATTGATA
GTTTATTGTG TACCAGTTTC TCCTGGTAAA TCTAGAATTG TTGCTCAGTT TACTCGCAAC
TTTGCCAAAA ACCTGCATTA TCTTATACCG CGTTGGTGGG AACACATCAA AATACGGAAT
CTAGTTCTAG ACGGAGATAT GATGCTGCTA CATCAGCAAG AATATTTATT GCAACAAAGA
CAAGAAAGCG AAAGTTGGAA AACTGCCTAT AAGTTGCCTA CAAGCGCAGA TCGTTTAGTA
ATTGAGTTTC GGACTTGGTT TGATAAATAT TGTCATGGTC AACTACCTTG GAGTAAGGTG
GGAATTAGTA ATCCAGAAAC TAAAATCAAT AACAACCGTG CTGTCATGTT GGATCGTTAC
CACCAACATA CCCAACATTG TAGTAGTTGC CGGAAGGCGC TGAAAAATCT ACAAAGATTA
CAAATCTTGC TTTTAACCTA TGTTGTAACT TCTGTTTGTG GAGTTGCAGT TCTTTCTGAT
GCTTTACGTA TGCAGCTAGG TCTACCAGTG GTCATTACAG CACTTTTAGG ATTGGGAGTT
TATTCTTGGT TGAAATTTTG GCTCATTCCT AAATTCTACT TTGTAGACTA TATCCATGCT
GAGAAATGA
 
Protein sequence
MQAEFNFFHH WYPLSPVEDL DSPRPVPVTL LGIRLVIWKP RDADNYRVFL DQCPHRLAPL 
SEGRVDDKTG NLMCSYHGWQ FDGRGICTHI PQAENPQLVA KNQQNFCVLS LPVRQENDLL
WVWPDAKSAE LAATTPLPLS PHIDTNKGFV WSSYIRDLEY DWQTLVENVA DPSHVPFSHH
GVQGNRDKAT SIPLNVVQST INLIEVSISK ALPTTITFQP PCLLEYAISI GDTDKKLGLI
VYCVPVSPGK SRIVAQFTRN FAKNLHYLIP RWWEHIKIRN LVLDGDMMLL HQQEYLLQQR
QESESWKTAY KLPTSADRLV IEFRTWFDKY CHGQLPWSKV GISNPETKIN NNRAVMLDRY
HQHTQHCSSC RKALKNLQRL QILLLTYVVT SVCGVAVLSD ALRMQLGLPV VITALLGLGV
YSWLKFWLIP KFYFVDYIHA EK