Gene Rcas_2141 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagRcas_2141 
Symbol 
ID5539621 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameRoseiflexus castenholzii DSM 13941 
KingdomBacteria 
Replicon accessionNC_009767 
Strand
Start bp2751333 
End bp2752553 
Gene Length1221 bp 
Protein Length406 aa 
Translation table11 
GC content62% 
IMG OID640894275 
Producthypothetical protein 
Protein accessionYP_001432244 
Protein GI156742115 
COG category[S] Function unknown 
COG ID[COG3214] Uncharacterized protein conserved in bacteria 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.129516 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones15 
Fosmid unclonability p-value0.124247 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGCCACTGA CGCTGACTCC CAATGCGATT CGTCCGTTGC TGCTGACCGT CCAGGGTCTC 
GACCGCCCAA AGAAACACCC GGCGACCAAA GACTCGGTTC TGGCAGCGAT GCAGCGGATG
AAGGCGCTTC AGATCGATAC GATCCATGTG GTCGCGCGTA GTCCATATCT GGTGTTATGG
AGCCGTCTTG GCGCATATGA GCCACGCTGG TTGACCGATC TGCTGGCGGA GCGGGCCATC
TTCGAATACT GGTCGCACGA GGCATGCTTT TTGCCCATCG AGGATTACCC GGCCTACCGC
TCGTTGATGC TGGCGGGGCA GACGCGCAGC AACACCTATG CGCGCCGGTG GCTGCACGAG
AACCGGACGA TTGCCGCAGC GTTGATGGAT CATATTCGCA ACAACGGACC GGTTCGCTCA
GCCGAATTCG CCCGTACAGA CGGCACGAGG GGAGGATGGT GGAACTGGAA GGTCGAAAAG
ATGGCGCTAG AAATGCTCTT CATCGTTGGC GATCTGATGA TCGACCGGCG TGAGCATTTT
CAGCGCCTCT ACGACCTACG CGAGCGAGTT TTGCCCGCAT GGGACGATAC CTGTGCACCC
GATGTGGAGG TAGCGCAGCG CACCCTGATC CTCGCAGCAG CACAGGCGCT CGGTGCAGCG
CCGGCGCGCT GGCTGGCAGA TTACTTTCGC ACCGGCAAGG CGGAAACGGC GCGCATTGCC
GCTGCTCTGG CAGCCGAAGG CGCCCTTGCG ATAGCGCATG TCGCAGGATG GCGCGAGCCG
GTCTACATTC ATCCACATCG CCTGCCGCTG GCGCAGGCTG CCGCCGATGG GGCGCTCCAA
TCAACAGTCA CCACGTTGCT TTCGCCATTC GATCCGGTCG TATGGGATCG GCGACGGGCG
CTGGAATTGT TTGGCTTCGA CTATCGCATC GAATGTTATA CTCCTGCATC CAAACGACGG
TATGGCTATT TTACGCTGCC CATCCTGCAC CGGGGGGCGC TGGTCGGGCG GCTCGACCCG
AAAGCGCACC GCAAAGACGG CATCTTCGAG GTCAAGGCGC TCTACCTCGA ACCAGGCGTC
GATCCCGACG AAGACCTGGC GATCAATCTG GCAGAGGCGT TGCGCTCCTG CGCCGTATGG
CACGGCACGC CGGAGGTCGT CGTCCGGTTC TGCGATCCGC CGGCGTTTGG CGCGTTGTTG
AAGCGCGCCT TGCGTCTCTG A
 
Protein sequence
MPLTLTPNAI RPLLLTVQGL DRPKKHPATK DSVLAAMQRM KALQIDTIHV VARSPYLVLW 
SRLGAYEPRW LTDLLAERAI FEYWSHEACF LPIEDYPAYR SLMLAGQTRS NTYARRWLHE
NRTIAAALMD HIRNNGPVRS AEFARTDGTR GGWWNWKVEK MALEMLFIVG DLMIDRREHF
QRLYDLRERV LPAWDDTCAP DVEVAQRTLI LAAAQALGAA PARWLADYFR TGKAETARIA
AALAAEGALA IAHVAGWREP VYIHPHRLPL AQAAADGALQ STVTTLLSPF DPVVWDRRRA
LELFGFDYRI ECYTPASKRR YGYFTLPILH RGALVGRLDP KAHRKDGIFE VKALYLEPGV
DPDEDLAINL AEALRSCAVW HGTPEVVVRF CDPPAFGALL KRALRL