Gene Rcas_0998 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagRcas_0998 
Symbol 
ID5538464 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameRoseiflexus castenholzii DSM 13941 
KingdomBacteria 
Replicon accessionNC_009767 
Strand
Start bp1303441 
End bp1304616 
Gene Length1176 bp 
Protein Length391 aa 
Translation table11 
GC content61% 
IMG OID640893141 
Productgalactokinase 
Protein accessionYP_001431124 
Protein GI156740995 
COG category[G] Carbohydrate transport and metabolism 
COG ID[COG0153] Galactokinase 
TIGRFAM ID[TIGR00131] galactokinase 


Plasmid Coverage information

Num covering plasmid clones14 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones35 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGCTTGATA CCGGAACGTT GCGCGCGCGC TTTCAGCAAC ACTACGGCAT ACATCCCTCT 
GTGATCGTTC GCGCGCCGGG GCGCGTCAAT CTCATCGGTG AGCATACCGA CTATAACGAT
GGTTTCGTTT TTCCGGTCGC CATTGATCGC GCCACCTACG TCGCGGCGCG TTTGCGCCAT
GATCAACTGG TGCGGGTGGC GTCGTCCGAC CTCAACGAAG AGGATACGTT CGCCATCGAT
CAGATCGAAC GCAGCAACCG ACCATGGCAC AATTACATTC GTGGCGTGGC GCTGGCGCTA
CGAGTTGCAG GGCATCCGCT TTTGGGGGCC GATCTCCTGA TCGCCAGTGA TGTCCCGCGC
GGTGCGGGGC TTTCGTCATC AGCCGCGCTC GAAGTCGCCG TCGGGTATGC GTTCCAGGTG
CTCAATAATC TGAACATTCT CGGCGAAGAA CTGGCATTGC TGGCACAGGG CGCAGAGAAC
AACTTCGTCG GCGTGCAATG CGGCATTATG GACCAGTTGA TTGCGGTGCT CGGTCGCGCC
GATCATGCGC TGCTGATCGA CTGTCGTGAT CTGTCCTATC GCGCCGTTCC GCTGCCCCCA
TCGGTTGCGG TCGTCATCTG CGACAGCCAT ATTCCGCGAA CTCTGGCGGC ATCGGCATAC
AACCAGCGGC GCCAGGAGTG CGATATGGCG GTTCAGTTGC TGCGCCGGTG GTATCCGGGT
ATTCGCGCAT TGCGCGATGT CAGCGAGGAT CACCTGGCAG CCCATTCCGA TGCGCTGCCA
GAGCCGATTC GCTCGCGCGC CCGGCATGTG GTCCGTGAAA ACCGTCGCAC ACTCCAGGGC
GCAGAAGCGC TCGAACGCGG CGATGTGGTC ACATTCGGGC GGTTGATGAA CGAGTCGCAC
GCCAGCCTGC GCGACGACTA TCAGGTGAGC CTGCCCGACA TCGACATTCT GGTCGAAACG
GCGCACCATC TGGCGGGATG TTACGGATCA CGCCTGACCG GCGCAGGATT TGGCGGGTGT
ACGGTGAGCC TGGTCGAGCG CAATGAAGTG GAATCGTTCA GCCGCGACCT GTTGCGCGTA
TATCACAATG CCACCGGTCG CACGGCCACC ATCTATGTAT GTCGCGCCAG CGATGGCGTT
GGGCGCGCCA CGGACAATGC AGGTCCACAG GAATGA
 
Protein sequence
MLDTGTLRAR FQQHYGIHPS VIVRAPGRVN LIGEHTDYND GFVFPVAIDR ATYVAARLRH 
DQLVRVASSD LNEEDTFAID QIERSNRPWH NYIRGVALAL RVAGHPLLGA DLLIASDVPR
GAGLSSSAAL EVAVGYAFQV LNNLNILGEE LALLAQGAEN NFVGVQCGIM DQLIAVLGRA
DHALLIDCRD LSYRAVPLPP SVAVVICDSH IPRTLAASAY NQRRQECDMA VQLLRRWYPG
IRALRDVSED HLAAHSDALP EPIRSRARHV VRENRRTLQG AEALERGDVV TFGRLMNESH
ASLRDDYQVS LPDIDILVET AHHLAGCYGS RLTGAGFGGC TVSLVERNEV ESFSRDLLRV
YHNATGRTAT IYVCRASDGV GRATDNAGPQ E