Gene NATL1_00651 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagNATL1_00651 
Symbol 
ID4779178 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameProchlorococcus marinus str. NATL1A 
KingdomBacteria 
Replicon accessionNC_008819 
Strand
Start bp71859 
End bp72950 
Gene Length1092 bp 
Protein Length363 aa 
Translation table11 
GC content33% 
IMG OID640083328 
Producthypothetical protein 
Protein accessionYP_001013894 
Protein GI124024778 
COG category 
COG ID 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.2704 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones15 
Fosmid unclonability p-value0.550522 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
TTGAATTTTA TAGATGATAA TGGTAATACT AAATTAAGTG TAGAAGATAA TAAAATCTTT 
GCACATTACG ATAATAAAAA ATTTGGATTA TGTGCTTATA CCATTTCAGT GCCTGGGTAT
TCAGATGATA TAGTCGGATG GGAGACAGAC AATAAAAAGA GTTGGCCTGG TATTTTTGTT
TATGATAATG GTGATTACTT TAAGTTTTAT GCTGAGGTTT TTAATAATAC TGGAATAATA
GTTTTTGAAG GTTTAGATAT TAATAATAAG GTTCAGAGAT ATCAAGTATA TGAGTTTTCG
TTAATAGATA TAGGTGTGAT TACAAAATAT TTTTCACAGA ATCCAGATGG TACTTTTTCT
AGTTATTTCT TTACTAGTTC AGAATATCCT GAGGATGAAT TTATATTAAA GCCACCTGAT
GGATCCTTGT CTTCCTGGGA GGCATTAGAT ACAGATCCTT ATTATCGGAC TCATGGTCAA
GGAACTTATA CAGATTTCTT AAAAGAATAT GGTCAAGTAA TTAATCCTGC TTCAGGGGAT
AGAAAATTTT TCCCTGCTAA AAGTTTTGAT TATCAATTCT TTGACTTAGG AAATCAAGAG
TATGGAATCA GACCAGATGC AGGTGGCACA GTAGATCCTT TAACAGGTTT CTCTTCAATT
CAGTTTATAG ATAAGTTTAT GACAATAGAA GCAGATATCA TAGGAGTCTT CGATCAAGTT
ACAGGGTTAA ATACAGACTC AGGCAAAATG TTTCGTCTTT ACAACGCTGC TTTTGCCCGG
TTCCCTGATG CTGATGGTTT GAGGTATTGG ATCAGTAATT TTAGTTCTGG GATTGATGAT
GAACGAGCAG TCTCATCTTC TTTTCTTGCC TCTGCTGAAT TTAAAGAACG TTATGGAGAA
AATATTACAC ATGAGACGTA TGTCGAAAAT CTCTATCTTA ATGTTTTGAA TAGGGAATTA
GATCAAGGCG GATACGATTA CTGGGTGGGT AATTTAAATA ATGGTATAGA GGAAAGACAT
GAAGTACTTT TAGGTTTTTC TGAGTCGGTA GAAAACAAAT TTCTTTTTAC TGAGATGACA
GGATTTGGTT GA
 
Protein sequence
MNFIDDNGNT KLSVEDNKIF AHYDNKKFGL CAYTISVPGY SDDIVGWETD NKKSWPGIFV 
YDNGDYFKFY AEVFNNTGII VFEGLDINNK VQRYQVYEFS LIDIGVITKY FSQNPDGTFS
SYFFTSSEYP EDEFILKPPD GSLSSWEALD TDPYYRTHGQ GTYTDFLKEY GQVINPASGD
RKFFPAKSFD YQFFDLGNQE YGIRPDAGGT VDPLTGFSSI QFIDKFMTIE ADIIGVFDQV
TGLNTDSGKM FRLYNAAFAR FPDADGLRYW ISNFSSGIDD ERAVSSSFLA SAEFKERYGE
NITHETYVEN LYLNVLNREL DQGGYDYWVG NLNNGIEERH EVLLGFSESV ENKFLFTEMT
GFG