Gene NATL1_00981 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagNATL1_00981 
Symbol 
ID4780682 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameProchlorococcus marinus str. NATL1A 
KingdomBacteria 
Replicon accessionNC_008819 
Strand
Start bp97602 
End bp98717 
Gene Length1116 bp 
Protein Length371 aa 
Translation table11 
GC content28% 
IMG OID640083361 
Producthypothetical protein 
Protein accessionYP_001013927 
Protein GI124024811 
COG category[R] General function prediction only 
COG ID[COG3380] Predicted NAD/FAD-dependent oxidoreductase 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.0926997 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones15 
Fosmid unclonability p-value0.536427 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGACTAATA TTTATGACTT TATAGTCATA GGTGCAGGAA TATCTGCTTG TACGTTTGCA 
TCACTTTTAA ATAAAAGATT CCCTGATCTT TCTATACTAT TAGTTGAACA TGGAAGGAGA
ATTGGTGGAA GAGCTACAAC GCGAAAATCA AGAAAAAATA GAATTCTTGA ATTTGACCAT
GGTTTGCCAT CTATTAATTT TAGAGAACGT ATTTCAGAAG ACATATTGGA GTTAGTTTCG
CCATTAATAA ATTCAGGAAA ATTGGTAGAT ATATCAAAGG ATATTTTATT AATCAATGAA
TTTGGCATTT TAAGTAATGC ATTTACTAAT GATATAATTT ATCGGAGTTC TCCTTTTATG
GCTAACTTCT GCGAGGAAAT AATTAATCAA TCTAATAACC CTAAAAAAAT AAATTTTTTA
TTTCAAACTC TAACTAAATC AATTAAGCGT ATAAATAACT TATGGGAGGT AAAAGTTAAT
ACTGGAAGAC ATATTAAATC TAAAAATCTG ATTTTATCCA GTTCTTTAAT AGCACATCCA
AGATGTTTGA ATCTTCTTCA AATTAATTCT TTACCACTTA GGGATGCCTT TATTCCAGGT
AAAGATAAAG TTGTTGATGC ATTAATAAAA GAAACAAGAA AATTAACTTA TATCATTAGA
AAAGTTTATA TTTTTCATGT TTCTAATTTG TCTTTATCTC AAAAATTTAA TTATCAATAT
TTACAAATTA TCTTTGCAAA TGTAATTAGA GAGGATTCAA ATTTTGAAAG AATTATTTTT
CAAAGACAAT CTGATGGATC TATCATTATA GCCTTACACT GTTCTTGTAT AAATAATCTG
CCCGAGATTA AGATTGATAA TATTACTAAA TCTTTGATAT CATTATTTGC AAACTATAAA
ATATTTTTGG ATTTGTTTTT ACAAGCAACC CTTATTGATA AAATGGATTG GAGAGCTTCT
CAACCTCTTA ATCACTTATT ATCCAAAGAA TTGCAGTGGT CTAATAGTAG TAAGATTGGT
TTTTGTGGAG ATTGGTTTGA TCTGAATTGC AGTGTAGGCG TAGAGTCTGC AATGAATAGT
TCACTCAGAC TGGTCAATTT TGTGAATCGG AATTGA
 
Protein sequence
MTNIYDFIVI GAGISACTFA SLLNKRFPDL SILLVEHGRR IGGRATTRKS RKNRILEFDH 
GLPSINFRER ISEDILELVS PLINSGKLVD ISKDILLINE FGILSNAFTN DIIYRSSPFM
ANFCEEIINQ SNNPKKINFL FQTLTKSIKR INNLWEVKVN TGRHIKSKNL ILSSSLIAHP
RCLNLLQINS LPLRDAFIPG KDKVVDALIK ETRKLTYIIR KVYIFHVSNL SLSQKFNYQY
LQIIFANVIR EDSNFERIIF QRQSDGSIII ALHCSCINNL PEIKIDNITK SLISLFANYK
IFLDLFLQAT LIDKMDWRAS QPLNHLLSKE LQWSNSSKIG FCGDWFDLNC SVGVESAMNS
SLRLVNFVNR N