Gene A9601_11991 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagA9601_11991 
SymbolbioB 
ID4717913 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameProchlorococcus marinus str. AS9601 
KingdomBacteria 
Replicon accessionNC_008816 
Strand
Start bp1015926 
End bp1016933 
Gene Length1008 bp 
Protein Length335 aa 
Translation table11 
GC content35% 
IMG OID640078915 
Productbiotin synthase 
Protein accessionYP_001009590 
Protein GI123968732 
COG category[H] Coenzyme transport and metabolism 
COG ID[COG0502] Biotin synthase and related enzymes 
TIGRFAM ID[TIGR00433] biotin synthetase 


Plasmid Coverage information

Num covering plasmid clones10 
Plasmid unclonability p-value0.496147 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGATAAATT CGAATAATCA GGTCTTAAGT GAAATTAGGT TCGATTGGAA TAAAAAGGAG 
ATATTAGAAA TTCTTAATAT GCCTCTGATC GATTTAATGT GGGAATCACA AATCGTTCAC
AGGAAATTCA ACAGCTATAA CATTCAATTA GCATCATTGT TCAGCGTAAA AACTGGTGGA
TGTGAGGAAA ATTGTTCATA TTGTAGCCAA TCAATTTATA GTGCTAGCGA AATAAAAAGT
CATCCACAAT TTCAAGTTGA AGAGGTTTTA GCAAGAGCTA AAATAGCAAA AAATGAGGGT
GCAGATAGGT TTTGTATGGG TTGGGCATGG AGAGAAATTA GAGATGGAAA ATCTTTTAAT
GCAATGTTAG AGATGGTTAG CGGTGTAAGA GATTTAGGAA TGGAAGCATG CGTTACTGCT
GGGATGCTTA CAGAAGAACA AGCTTCCCGA CTGGCTGATG CAGGTTTAAC AGCGTACAAC
CACAATCTTG ATACTAGCCC TGAGCATTAT AAAAATATTA TTACGACAAG AACTTATCAA
GACAGACTAG ATACTATCAA AAGAGTGAGA AATGCAGGAA TAAATGTTTG TTGTGGAGGA
ATAATAGGCT TGGGTGAGAC TAATGGCGAT AGAGCATCTC TTTTGGAAGT GCTTTCAAAC
ATGAATCCAC ACCCTGAAAG TGTTCCAATA AATTCATTAG TAGCTATTGA AGGTACTGGT
TTAGAAGATA ATCAAGAAAT TGATTCTATT GAGATGATAA GGATGATAGC TACAGCCAGA
ATTCTTATGC CTAAAAGTAA AATAAGATTA AGTGCAGGAA GAGAAAAGCT CTCAAAAGAA
GCCCAAATCT TATGTTTTCA ATGCGGTGCA AATTCCATTT TTTACGGAGA TGAGTTACTC
ACTACTTCAA ATCCATCTTT TCAGTCAGAC AGAAAACTTC TCAAAGAGGT TGGAGTATCA
TTTAACAAAG ATTTTGAAAC TTGTGAAAAA ACATTATCTT CTTTATGA
 
Protein sequence
MINSNNQVLS EIRFDWNKKE ILEILNMPLI DLMWESQIVH RKFNSYNIQL ASLFSVKTGG 
CEENCSYCSQ SIYSASEIKS HPQFQVEEVL ARAKIAKNEG ADRFCMGWAW REIRDGKSFN
AMLEMVSGVR DLGMEACVTA GMLTEEQASR LADAGLTAYN HNLDTSPEHY KNIITTRTYQ
DRLDTIKRVR NAGINVCCGG IIGLGETNGD RASLLEVLSN MNPHPESVPI NSLVAIEGTG
LEDNQEIDSI EMIRMIATAR ILMPKSKIRL SAGREKLSKE AQILCFQCGA NSIFYGDELL
TTSNPSFQSD RKLLKEVGVS FNKDFETCEK TLSSL