Gene PICST_34605 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagPICST_34605 
SymbolTCD4 
ID4851794 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameScheffersomyces stipitis CBS 6054 
KingdomEukaryota 
Replicon accessionNC_009068 
Strand
Start bp2846132 
End bp2847289 
Gene Length1158 bp 
Protein Length385 aa 
Translation table 
GC content45% 
IMG OID640393502 
ProductFe(II)-dependent sulfonate/alpha-ketoglutarate dioxygenase-like protein 
Protein accessionXP_001387108 
Protein GI126275620 
COG category[Q] Secondary metabolites biosynthesis, transport and catabolism 
COG ID[COG2175] Probable taurine catabolism dioxygenase 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones23 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones16 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGTCTTCAA CTGCAACTTA CGGAAACTTC GACACCCATT TCTTTGCTGG CCAAGACGAA 
ATTGGGGAAG ACGGAATTTT GACCATCAAC AAGAGAAACA GAGAGCAATC TCATTACCCT
GAATTCTTGC CTACTTGGGA TCCCAGCCAG AAGTACCCAC CTTTGAAGTT TTTCAAGCAT
GAAGATCCTG GAAAGAGAGC CGATCCATCT TTCCCCAATC TATTTGCCAA AGATCACGAA
CAAATCGTCA AGAAAGTCAC TCCCAAGTTG GGATCCGAAG TTAGGGGAGT TCAATTGTCT
CAATTGGACT CAGCTGGTAA AGACGAGTTG GCTCTTTTTG TGGCACAAAG AGGAGTGGTA
ATCTTCCGCG ACCAGGATTT CGCAGCCAAA GGTCCAGCTT TCGCAGTTGA ATATGGTAAA
CACTTCGGAA GATTGCACAT CCACCCAACA TCTGGTGCTC CAAGAAACCA CCCAGAGTTG
CACATCACCT ACAGAAGAGC TGATCCCGGC GAATTCGAGA GAGTTTTCTC CAATAGCACT
AATGCTGTTC AGTACCATAC TGATGTATCC TACGAGTTGC AACCAGCAGG GATCACTTTT
TTCTCAGTAT TGGAAGGGCC GGAATCCGGT GGTGACACCA TCTTCGCCGA TTCAGTCGAA
GCATACAACA GATTATCTCC AGCTTTCCAG AAGAGGTTGG CCGGCTTACA TGTGTTGCAT
ACTTCCGAAG ACCAAGCTTC TAACTCTAGA GGCCAAGGTG GAATTGAAAG AAGAAAGCCA
GTTTCAAACA TCCATCCATT GGTCAGAATT CATCCAGTTA CCGGTGCAAA GAGTTTGTTT
GTCAATAGAT CATTTGCTAG AAGAATCGTT GAGTTGAAAG AAGAAGAATC CGAGTCTTTG
CTTAAATTCT TGTACGACCA CATTGAGCAA TCCCATGACT TGCAATTGAG AGCCAATTGG
GAACCAAACA CAGTGGTTAT CTGGGACAAT AGAAGGGTGC ACCACTCAGC CATCATCGAC
TGGGAAACTG CAGTTTCTAG ACATGCCTTC AGAATCACTC CACAAGCCGA AAGACCTGTG
GAAGACTTGA AGGACTTGAA TAAAGAAGAG TACGACGTTG GTGATGTTGC TGAAGCTTTG
AAAGCTGTTT TACATTAG
 
Protein sequence
MSSTATYGNF DTHFFAGQDE IGEDGILTIN KRNREQSHYP EFLPTWDPSQ KYPPLKFFKH 
EDPGKRADPS FPNLFAKDHE QIVKKVTPKL GSEVRGVQLS QLDSAGKDEL ALFVAQRGVV
IFRDQDFAAK GPAFAVEYGK HFGRLHIHPT SGAPRNHPEL HITYRRADPG EFERVFSNST
NAVQYHTDVS YELQPAGITF FSVLEGPESG GDTIFADSVE AYNRLSPAFQ KRLAGLHVLH
TSEDQASNSR GQGGIERRKP VSNIHPLVRI HPVTGAKSLF VNRSFARRIV ELKEEESESL
LKFLYDHIEQ SHDLQLRANW EPNTVVIWDN RRVHHSAIID WETAVSRHAF RITPQAERPV
EDLKDLNKEE YDVGDVAEAL KAVLH