Gene Ssol_1080 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagSsol_1080 
Symbol 
ID
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameSulfolobus solfataricus 98/2 
KingdomArchaea 
Replicon accessionCP001800 
Strand
Start bp1011963 
End bp1013168 
Gene Length1206 bp 
Protein Length401 aa 
Translation table11 
GC content35% 
IMG OID 
Productputative transcriptional regulator, GntR family 
Protein accessionACX91323 
Protein GI261601720 
COG category 
COG ID 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones14 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGTTTGAGA GATTTTTATC CACCGAAACT AAGTATTTAC GTACCTCAGA AATAAGAGAT 
CTACTAAAAC TAACGGAAGG TAAGAATGTA ATTAGCCTTG CGGGAGGTTT ACCTGATCCT
CAGACTTTTC CAGTAGAAGA AATTAAAAAA ATAGCAGATG ACATCTTATT GAATAGTGCT
GATAAGGCAT TACAATATAC TGCAACTGCT GGTATATCTG AATTTAGAAG AGAATTGGTG
AACTTATCTA GATTAAGGGG AATTAGTGGA ATAGATGAAA GAAACGTCTT TGTTACAGTA
GGAAGTCAAG AAGCACTTTT CATGATATTT AATATATTAC TAGATCCGGG AGACAATGTA
GTAGTCGAGG CGCCAACTTA TTTAGCAGCT TTAAATGCCA TGAGAACTAG AAAGCCAAAT
TTCATATCGA TAACAGTAAC GGAAATGGGC CCAGATCTAG ATGAATTGGA GAGAAAAATA
AAAGATGCCC ATAGCAATGG GAAGAAGGTT AAACTGATGT ATGTGATTCC AACAGCCCAG
AATCCGGCGG GTACTACGAT GAATACAGAG GATAGGAAAA GACTTTTAGA GATTGCATCG
AAATATGATT TCTTAATTTT TGAGGATGAT GCTTATGGGT TCTTAGTATT CGAGGGAGAA
AGTCCACCAC CAATTAAAGC CTTCGATAAA GAAGGAAGAG TAATTTATAC TAGCACATTT
AGTAAAATAC TTGCACCCGG TTTAAGGTTA GGATGGGTAA TTGCTCATGA AGATTTCATA
AAGGAAATGG AACTATATAA ACAAAATGTT GATTTGCATA CACCTTCATT ATCACAATAT
ATTGCAATGG AGGCTATAAG GAGGGGTATA ATTCAAAATA ATTTACCTAA GATAAGGAGA
GTGTATAAGG AAAAAAGAGA TGTAATGCTA GAGGCTATTG AAACTTATTT CCCTAATGAT
GCCAGATGGA CTAAACCAGT TGGTGGAATG TTTGTTTTTG CTTGGTTGCC ACAAAAAATA
GATACTACTA AGATGTTAGA AAAAGCCTTA CAAAGGGGTG TAGCTTATGT ACCAGGTTCT
AGTTTCTATG CTGACTATAG TGGAAAGAAT ACTATGAGGA TCAACTTTAG TTTTCCTAAG
AAAGAAGAAT TAATAGAGGG AATTAAGAGG CTAGGAGATA CGATAAAGCA TGAGCTCTCT
ACTTAA
 
Protein sequence
MFERFLSTET KYLRTSEIRD LLKLTEGKNV ISLAGGLPDP QTFPVEEIKK IADDILLNSA 
DKALQYTATA GISEFRRELV NLSRLRGISG IDERNVFVTV GSQEALFMIF NILLDPGDNV
VVEAPTYLAA LNAMRTRKPN FISITVTEMG PDLDELERKI KDAHSNGKKV KLMYVIPTAQ
NPAGTTMNTE DRKRLLEIAS KYDFLIFEDD AYGFLVFEGE SPPPIKAFDK EGRVIYTSTF
SKILAPGLRL GWVIAHEDFI KEMELYKQNV DLHTPSLSQY IAMEAIRRGI IQNNLPKIRR
VYKEKRDVML EAIETYFPND ARWTKPVGGM FVFAWLPQKI DTTKMLEKAL QRGVAYVPGS
SFYADYSGKN TMRINFSFPK KEELIEGIKR LGDTIKHELS T