Gene NATL1_08361 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagNATL1_08361 
SymbolhisA 
ID4780472 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameProchlorococcus marinus str. NATL1A 
KingdomBacteria 
Replicon accessionNC_008819 
Strand
Start bp768731 
End bp769501 
Gene Length771 bp 
Protein Length256 aa 
Translation table11 
GC content36% 
IMG OID640084111 
Product1-(5-phosphoribosyl)-5-[(5- phosphoribosylamino)methylideneamino] imidazole-4-carboxamide isomerase 
Protein accessionYP_001014659 
Protein GI124025543 
COG category[E] Amino acid transport and metabolism 
COG ID[COG0106] Phosphoribosylformimino-5-aminoimidazole carboxamide ribonucleotide (ProFAR) isomerase 
TIGRFAM ID[TIGR00007] phosphoribosylformimino-5-aminoimidazole carboxamide ribotide isomerase 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.103566 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones11 
Fosmid unclonability p-value0.0927273 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGGAAATCA TACCTGCAAT AGATCTACTG AATGGTAAAT GTGTTCGACT AAATCAAGGA 
AATTATAATG AAGTTACTAA GTTCAATAGT GATCCTGTAA AACAAGCACA AATTTGGGAA
AGCAAGGGAG CAAAACGACT ACATCTTGTA GATCTTGATG GTGCTAAGAC AGGTGAGCCC
ATAAATGATC TAACTATAAA AGAGATAAAA AAATCTATTA CAATACCTAT TCAACTTGGC
GGTGGAATTA GGAGTATTGA TCGTGCGAAA GAATTATTCG ACATTGGAAT AGACAGAATT
ATTTTAGGAA CAATTGCAAT AGAGAAGCCC GAATTAGTTA AAGACCTATC TAAAGAATAT
CCAAAAAGAG TTGCAGTAGG AATTGATGCC AAAGAGGGAA TGGTAGCCAC TCGAGGTTGG
TTAAAACAAA GCAAAATATC TTCTCTAGAC TTAGCAAAAC AACTTAACGA TCTTGACTTA
GCGGCAATCA TATCAACTGA CATTGCTACC GATGGCACTC TAAAAGGACC TAATGTTCAA
GCCTTGAGAG AAATAGCTGA GATAAGTATT AATCCAGTAA TTGCCTCAGG GGGGATAGGT
TCAATAGCTG ATTTAATTTC ACTAGCAGAT TTCGCGGATG AAGGTATTGA AGGAATAATC
GTAGGCAGAG CCCTATATGA CGGCTCAATA GATTTAAAGG AAGCGATTTT AACTCTAAAA
AATCTTCTTC TGCAAGATGC TTTCAATGAG AAAGATAAAT TTCTTGTCTA A
 
Protein sequence
MEIIPAIDLL NGKCVRLNQG NYNEVTKFNS DPVKQAQIWE SKGAKRLHLV DLDGAKTGEP 
INDLTIKEIK KSITIPIQLG GGIRSIDRAK ELFDIGIDRI ILGTIAIEKP ELVKDLSKEY
PKRVAVGIDA KEGMVATRGW LKQSKISSLD LAKQLNDLDL AAIISTDIAT DGTLKGPNVQ
ALREIAEISI NPVIASGGIG SIADLISLAD FADEGIEGII VGRALYDGSI DLKEAILTLK
NLLLQDAFNE KDKFLV