Gene Aazo_4151 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagAazo_4151 
Symbol 
ID9341956 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism name'Nostoc azollae' 0708 
KingdomBacteria 
Replicon accessionNC_014248 
Strand
Start bp4224666 
End bp4225796 
Gene Length1131 bp 
Protein Length376 aa 
Translation table11 
GC content42% 
IMG OID 
Productglutamate 5-kinase 
Protein accessionYP_003722708 
Protein GI298492531 
COG category 
COG ID 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones17 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGCCCCATG TCCCATGCCC CATGACTAAA ACAATAGTTG TTAAAATTGG TACTTCTAGC 
CTTACTCAAC GAGAAACTGG ACAACTAGCC CTTTCCACCA TTGCTACCTT AACGGAAACC
CTGTGCAATT TAAGACTCCA GGGTCATCGC GTAATTTTGG TTTCTTCCGG TGCTGTGGGT
GTGGGTTGTG CGTGTTTAGG TTTAACAGAA CGTCCCAAAG TGATCGCTCT CAAACAAGCG
GTAGCAGCTG TTGGACAAGG TAGGCTAATA CGTATCTATG ATGATTTATT TACTACTTTA
CAACAACCTA TAGCCCAAGT ATTATTAACA CGCGCTGATT TGGTACAACG TAGCCGCTAT
CTAAATGCTT ACAATACTTT TCAGGAATTG CTACGACTAG GAGTAATTCC GGTAGTGAAT
GAAAATGATA CTGTGGCTGT AGAGGAATTG AAATTTGGTG ATAACGACAC CCTTTCTGCT
TTAGTTGCCA GTTTAGTGGA AGCGGATTGG TTATTTTTAC TGACAGACGT TGAGAAATTA
TATTCTGCTG ATCCTCGTTC TGTACCTGAT GCCCGTCCTA TCAGTTTGGT AAGTAATATG
AGGGAATTGG CAGATTTGCA AATTCAAACC GGGGGACAGG GTTCTCAGTG GGGTACTGGT
GGAATGGTAA CAAAAATATC TGCTGCCAGA ATTGCGATCG CAGCGGGTGT GCGAACTATA
ATTACTCAAG GGCGTTTTCC TCACAATATT GAGAAAATTA TCCAAGGGGA AGCTATAGGA
ACGCATTTTG AACCGCAACC AGAACCAACC TCAGCTAGAA AACGCTGGAT AGCTTATGGT
TTAGTACCGA TGGGTAAATT ATATTTAGAT GATGGGGCTA TTAATGCTAT TTCCCAAGCA
GGAAAATCTC TGTTGGCTGC GGGAATTAAA GCTGTACAAG GGGAATTTGA CCATCAGGAA
GCGGTACAAT TGTGCGATGG CACAGGTAAT GAAATTGCCA GAGGTTTGGT GAATTATAAC
AGTGAAGAAT TACAAAAAAT TTGTGGTTGT CATTCACGGG ACATTGCGGG AATTTTGGGT
TATGCAGGTG CGGAAACTGT AATTCATCGG GATAATTTGG TGTTGATTTA G
 
Protein sequence
MPHVPCPMTK TIVVKIGTSS LTQRETGQLA LSTIATLTET LCNLRLQGHR VILVSSGAVG 
VGCACLGLTE RPKVIALKQA VAAVGQGRLI RIYDDLFTTL QQPIAQVLLT RADLVQRSRY
LNAYNTFQEL LRLGVIPVVN ENDTVAVEEL KFGDNDTLSA LVASLVEADW LFLLTDVEKL
YSADPRSVPD ARPISLVSNM RELADLQIQT GGQGSQWGTG GMVTKISAAR IAIAAGVRTI
ITQGRFPHNI EKIIQGEAIG THFEPQPEPT SARKRWIAYG LVPMGKLYLD DGAINAISQA
GKSLLAAGIK AVQGEFDHQE AVQLCDGTGN EIARGLVNYN SEELQKICGC HSRDIAGILG
YAGAETVIHR DNLVLI