Gene Aazo_4719 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagAazo_4719 
Symbol 
ID9342526 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism name'Nostoc azollae' 0708 
KingdomBacteria 
Replicon accessionNC_014248 
Strand
Start bp4824541 
End bp4825680 
Gene Length1140 bp 
Protein Length379 aa 
Translation table11 
GC content44% 
IMG OID 
Producthypothetical protein 
Protein accessionYP_003723040 
Protein GI298492863 
COG category 
COG ID 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones15 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGACAGACC AACCTCCCGT AGCTAACCCA ATGAATGCCG CAGCTATCCC TATGAATCGA 
GTAGCTCCAA CCCCTATAAA CAACAATGTT AATAAGCCAG TTAGTGTTTC GGGGAAAAAT
ATTCTCAGCG TGGATTTAGG TAGAACCTCT ACCAAAACTT GTGTCAACCG CGAACCTGCC
AATGTTGCTT TCATTCCTGC CAATGTTAAG CAAATGTCAA TAGAACAAAT ACGCGGTGGT
GTGTTTGAAT CTAAAGCCAC AGATCCTTTG ATGGATTTGT GGATGGAATA TCAAGGTAAC
GGATATGCTG TGGGTCAACT GGCAGCAGAT TTTGGTGCTA ATTTAGGAGT AGGCCAATCT
AAGGTAGAAG ATGCACTAGC TAAAGTTCTG GTAGCCGCTG GTTACTTTAA GCTCAAAGAT
GAAATTTCTG TCATCGTTGG TTTGCCTTTC CTTTCTCTAG AGCAATTTGA ACGGGAAAAA
GCTCAACTGA TGAGTTTGAT TAGTGGTCCC CATGTGATGA ATTTCCGTGG CGAAACTGTA
TCGTTAAACG TCACTAAAGT TTGGGTAATG CCAGAAGGCT ATGGTAGCTT GCTGTGGTGT
GAAACCCAAC CAAATAAAGG CTCTTCGATG CCTGATTTGA CAAAGGTATC AGTGGGTATA
GTGGATATTG GACATCAAAC CATTGACTTG TTGATGGTTG ATAACTTCCG CTTTGCTAGA
GGTGCTTCCA AGAGTGAAGA CTTTGGTATG AGCAAGTTTT ATGAAATGGT GGCTAAGGAA
ATAGAAGGTG CTGATAGTCA ATCTCTAGCA CTGATTTCTG CTGTAAACAA GCCCAAAGGT
GATCGCTTTT ATCGTCCTAA AGGTGCAAGC AAACCCGCTA ACCTAGACGA TTTTCTCCCT
AACCTCACAG AGCAGTTCTC ACGGGAAATT TGCTCGATCG TGTTAGCCTG GTTACCAGAG
CGTGTAACTG ATGTGATTAT CACAGGTGGC GGTGGTGAGT TCTTCTGGGA AGATGTACAA
CGTCTGCTCA AAGAAGCCAA AATTCATGCC CATTTGGCGG CACCCTCACG CCAAGCTAAT
GCTTTAGGGC AGTATATTTA TGGAGAGGCG CAATTATCTG CTGTTCGCGC TGCTAGGTAA
 
Protein sequence
MTDQPPVANP MNAAAIPMNR VAPTPINNNV NKPVSVSGKN ILSVDLGRTS TKTCVNREPA 
NVAFIPANVK QMSIEQIRGG VFESKATDPL MDLWMEYQGN GYAVGQLAAD FGANLGVGQS
KVEDALAKVL VAAGYFKLKD EISVIVGLPF LSLEQFEREK AQLMSLISGP HVMNFRGETV
SLNVTKVWVM PEGYGSLLWC ETQPNKGSSM PDLTKVSVGI VDIGHQTIDL LMVDNFRFAR
GASKSEDFGM SKFYEMVAKE IEGADSQSLA LISAVNKPKG DRFYRPKGAS KPANLDDFLP
NLTEQFSREI CSIVLAWLPE RVTDVIITGG GGEFFWEDVQ RLLKEAKIHA HLAAPSRQAN
ALGQYIYGEA QLSAVRAAR