Gene Rcas_2666 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagRcas_2666 
Symbol 
ID5540148 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameRoseiflexus castenholzii DSM 13941 
KingdomBacteria 
Replicon accessionNC_009767 
Strand
Start bp3437135 
End bp3438253 
Gene Length1119 bp 
Protein Length372 aa 
Translation table11 
GC content58% 
IMG OID640894788 
ProductABC-type sugar transport system periplasmic component-like protein 
Protein accessionYP_001432755 
Protein GI156742626 
COG category[G] Carbohydrate transport and metabolism 
COG ID[COG1879] ABC-type sugar transport system, periplasmic component 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.157111 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones26 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGAAACGCC GACAATTCGC ACTGTTTGCG ATCACGCTCA TGCTCGCCAC CCTGATTGCG 
GCATGCGGCG GAACGCCCAC CACCACCGCA CCGACCGCTG CGCCAGCGCC ACAACCCACG
ACGGCGCCAG CGCCCGAGGG CAGAAAGTTC ACAATCGGCA TCTCCAACCC GTTCATCAGC
AGCGAATACC GCACCCAGAT GATCCAGTCG TTGATCGAGG TCAATAAGGA ATACATGGAA
CGGGGGATCA CCAACGAACT CGTGATCGAG AGCGCCGATA CCGATGTCGC CGGTCAGATC
CAGCAATTGC AGAATCTCAT TAACAAAGGG GTCGATGCCA TCCTGGTGAA CCCCAGCGAT
GTCAATGGTC TTAACGACAC GCTTCAGGAA GCGATCAACA AGGGGATCAT CGTCATCTCG
GTCGATCAGG AACTGAACAC CCCCGGCGTC TACAACGTCG GCATCGATCA GAAGGAATGG
GCGAAGATTT CAGCCCGCTG GCTGGCGGAG AAGCTTGGCG GACAGGGAAA TATTGTGCTG
ATCGAGGGCT TCCCCGGGCA TCCGGCGAAC GTTGCGCGCA TGGAGGGCGT CGAGGAAGTG
CTGAAGGAGT ATCCCAATAT CAAGGTGCTA GGGCGTGAAA CCGGCAAGTG GGACGAAGCC
ACCGGTCAGC AGGTGATGTC GAACTTCCTG GCGTCGTTCC CCAATCTCGA CGGCTACTGG
ACGCAGGATG GCATGGCGAT TGGCGCGATG CAGGCGGTGA TGGCCGCCAA CCCGTCGAAG
TGGCCCGTGC TCGTCGGCGA GGGGCGCTGC CAGTTCTTGC AGTTGTGGGA TCAGCGCCTG
AAGGAAGACC CCAACTTCGA GACGATTGCC GTCGCCAATC CGCCCGGCGT ATCGCCAACC
GGTCTGCGCA TTGCCATCAA TATGCTTCAG GGCAAGCAGG TGGACAAGAG TAAACTTGGA
GGGGCGAATG GACTGTCGTT CGTCATTCCG GTGCCGGTGA TTGTGACGAA AGACAATTTC
CAAGAGGTGT TCACCACAGT GTGCAAGGAT AAGCCGGCCA CCTACCTGCT CGACGGCATT
ATGACCGACG AGGAAGTGCA GCAGTTCTTC GTGAAGTAG
 
Protein sequence
MKRRQFALFA ITLMLATLIA ACGGTPTTTA PTAAPAPQPT TAPAPEGRKF TIGISNPFIS 
SEYRTQMIQS LIEVNKEYME RGITNELVIE SADTDVAGQI QQLQNLINKG VDAILVNPSD
VNGLNDTLQE AINKGIIVIS VDQELNTPGV YNVGIDQKEW AKISARWLAE KLGGQGNIVL
IEGFPGHPAN VARMEGVEEV LKEYPNIKVL GRETGKWDEA TGQQVMSNFL ASFPNLDGYW
TQDGMAIGAM QAVMAANPSK WPVLVGEGRC QFLQLWDQRL KEDPNFETIA VANPPGVSPT
GLRIAINMLQ GKQVDKSKLG GANGLSFVIP VPVIVTKDNF QEVFTTVCKD KPATYLLDGI
MTDEEVQQFF VK