Gene Rsph17025_3891 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagRsph17025_3891 
Symbol 
ID5085439 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameRhodobacter sphaeroides ATCC 17025 
KingdomBacteria 
Replicon accessionNC_009429 
Strand
Start bp789027 
End bp790682 
Gene Length1656 bp 
Protein Length551 aa 
Translation table11 
GC content67% 
IMG OID640485450 
Producthypothetical protein 
Protein accessionYP_001170051 
Protein GI146279893 
COG category[N] Cell motility
[T] Signal transduction mechanisms 
COG ID[COG0840] Methyl-accepting chemotaxis protein 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones17 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones18 
Fosmid unclonability p-value0.452082 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGAAACTGA ACATCAAGAT CAAGCTGGCC GGTGCCTTCT TCCTGGTCTT CCTGCTCATG 
GGAACCGGCA CGATACTGGG AATCATCGAT CTGCGGCACT CGAACCAGGT GCTCCAGACG
ATCGTCGAGA AACAGGCCGC GCGCGTCGAG TCAGCGAGCC GGCTGGAGAT CCAGCAGACA
CAGTTCAACG TCGTCCTGCG GGACTATGTG GTCGCCGAGG ATGAGGCCAA ACGCGCCGCG
CTCAAGCAGG ACATCGTGCA GATCCGCGCC GACATGAGCG CAAGCATCGA GCGGCTCGAG
GCGCTGGCCG ACGATGTCGG GATGCCGATG ATCAAGGCCT ATGCCGAGCA GCGCAAGGCC
GCCGCCGCGA TCAACAACCG CGTGTTCGAG CTTGCCGATG GCGGCGATAC GGCGGCCGCC
TCGCGCCTGC TTGCCGGCGA GGCGCGTCAG GGAATGGAGA AGCTCGCCGC CAATCTCGAG
GCCTTCCGCA GCACCTACAG CAACCAGATG GCTGAGGCGA CCGCCTCTGC CAGCCGCGAT
CTGGCGGCGA GCGTGCTCAA CCTGTCCCTC CTCGCTCTGG CCGGGGTGCT GTTCGGGACA
ATCGCCGCCA CGCTCGTGAC GGCCTCGATC GCGCGGGGCC TTCAGCGGAC ACTGGACCTG
ACGCAACGCG TTGCGCACGG TGATCTGACG ACGCTGGCCG ATGACCGGGG CTCGGACGAG
ATCGCGCAAC TGCTCAAGGC GAGCAACTCG ATGATCCTCC GCCTCCGCGA GGTCGTGGGC
CGCGTCACGC TTGCGACCAA TCAGGTGGCC GCCAACAGCC GGATCATGGC CTCCACGTCC
GAGCAGCTGT CGCAGGGCAG CAGCGAACAG GCCTCGTCCA CGGAAGAAGC GTCCGCCTCG
GTCGAGCAGA TGGCGGCCAA CATCAAGCAG ACGGCGGACA ACTCCGCCCG CACGGAACAG
ATTGCCATCA AATCGGCCGA GGATGCGCGT GCCTCCGGTG GCGCGGTGCG CGAGGCGGTG
TCGGCCATGG GTGCGATTGC CGAACGTATC CTCGTGGTGC AGGAGATCGC CCGTCAGACC
GATCTTCTGG CGCTCAACGC GGCTGTCGAG GCCGCCCGCG CAGGCGAGCA CGGACGCGGC
TTTGCCGTCG TCGCGGCCGA GGTGCGCAAG CTGGCCGAGC GGAGCCAGTC CGCGGCCGCG
GAAATCTCGC AGCTTTCGTC GCGGACCTCT GCAGCGGCTT CCACCGCGGG CGAAATGCTC
GAGCGGCTGG TGCCTGACAT CGAGCGCACC TCCACGCTTG TCTCCTCGAT CTCGGTCGCC
TCGCGCGAAC TTTCCACGGG GGCACAACAG GTGGCACTGG CGATCCAGCA GCTGGATCAG
GTGACCCAGC AGAACAGCCA CTCGGCAGAG GCTCTGGCCG AGGGGGCAGG AGAGCTGTCG
ACCGAAGCCG ACCAGCTCAA GGACGCGGTC GGGTTCTTCC TGATCGACGC GACGCCGGAG
CGGCCGGCCC GTGATCCTCA GCCCGCGCCG AAGTCCGCCC CGCCGGCCGT GCGCAAGCCT
GCCCTTCAAG TCGCCGCAAA ACCGAAGGGG TTCCACTTCG ACATCGGGGA GAGCGACATG
GACGAACTGG ACGCAGCGTT CCAGCGCACC GCCTGA
 
Protein sequence
MKLNIKIKLA GAFFLVFLLM GTGTILGIID LRHSNQVLQT IVEKQAARVE SASRLEIQQT 
QFNVVLRDYV VAEDEAKRAA LKQDIVQIRA DMSASIERLE ALADDVGMPM IKAYAEQRKA
AAAINNRVFE LADGGDTAAA SRLLAGEARQ GMEKLAANLE AFRSTYSNQM AEATASASRD
LAASVLNLSL LALAGVLFGT IAATLVTASI ARGLQRTLDL TQRVAHGDLT TLADDRGSDE
IAQLLKASNS MILRLREVVG RVTLATNQVA ANSRIMASTS EQLSQGSSEQ ASSTEEASAS
VEQMAANIKQ TADNSARTEQ IAIKSAEDAR ASGGAVREAV SAMGAIAERI LVVQEIARQT
DLLALNAAVE AARAGEHGRG FAVVAAEVRK LAERSQSAAA EISQLSSRTS AAASTAGEML
ERLVPDIERT STLVSSISVA SRELSTGAQQ VALAIQQLDQ VTQQNSHSAE ALAEGAGELS
TEADQLKDAV GFFLIDATPE RPARDPQPAP KSAPPAVRKP ALQVAAKPKG FHFDIGESDM
DELDAAFQRT A