Gene RPB_0761 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagRPB_0761 
Symbol 
ID3909249 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameRhodopseudomonas palustris HaA2 
KingdomBacteria 
Replicon accessionNC_007778 
Strand
Start bp854506 
End bp856185 
Gene Length1680 bp 
Protein Length559 aa 
Translation table11 
GC content66% 
IMG OID637882653 
Productflagellar hook-associated protein 
Protein accessionYP_484383 
Protein GI86747887 
COG category[N] Cell motility 
COG ID[COG1345] Flagellar capping protein 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones14 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones14 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGGCCACCG TCACCAGTTC GACTTCCGCC TCGGCGACGG CCGCGCTCGC GACCACGACG 
TCCGGCACCA CGACCACGAC GAGCGTCGAT TGGGATGCGT TGATCGAAGC GCAGGTGGCG
ACCAAGACCG CCGCCGCCGA CACCATCGAG ACCAGCATCA CGGCCAACGA AGCCAAGATC
TCGGCCTACC AAAATCTGCA GACGCTGCTC GACACGCTCG TGACCAGCAC CACCTCGCTG
TCGAAGTCGA TCGTCAACTC GCTGTCCGAC AGCACCTTCG GCGCCCGCGC GGCCACGATC
ACCTCGAGCG GCGACGTCAG CGCCAGCTCC GCGGTGTCGA TGTCGATCAG CAACGGCGCC
GCCACCGGCG ACCATACGCT CACGGTCGAG CAGATCGCCA CCGCGCACCG CGTGATCGGA
ACCAGCGTCG CGGACAAATC CGCGGATATG GGCCTGACCG GCGTGTTCTC GCTCGGCCTG
GCCGGCGGCA CCAGCGTCGA CGTCTCGATC ACCAGCGGCA TGTCGATGGA AGACATCGCC
GACACCATCA ATGCGCAGAG CGACAGCACC AACGTCCAGG CCTCGATCAT CCAGATCTCG
AGCACCGAAT ACGCGCTGAC GCTGACCGCG CTGAACGACA ACGCCGAGAT CACCACCAGC
GTCGTCTCCG GCGACGACGT GCTGACGACG CTCGGCGTCA CCGATTCCGC CGGCGACTTC
ACCGACGTGC TGCAGGAGCC GCAGCCGGCG CTGTTCACGG TCGACGGCAT CTCGCTGACC
CGCGACACCA ACGACATCAC CGACGTGCTG AGCGGCGTGA CCTTCAGCCT GCTGCAGGCG
ACGCCGGACG GCTCGACCAT CAATCTCAGC ATCGACGTCG ACGCCGACCA GATCGCGGCC
GCGCTGGAGG AGTTCGTCAC CGCCTACAAC GCCGTCCGCG AGGAGGTCAC CGCGCAGCAG
ACGCTGACCT CGGACGGGAC CGCGGATTCC AGCGCCGTGC TGTTCGGCGA CGGCACCATG
CGCAGCATCA TGACGCAGAT CGAACAGGCG ATGAACTCCA CCGTCGGCGG ACTGTCGATG
ACCGACCTCG GGCTGTCGTT CACCGACACC AATACGCTCG AGTTCGACAC CAGCGTGCTG
TCGGCCACGC TGACCGAAGA CCTCTCGGGC GTGATCGCTC TGCTGGCGTC GAAGACGACG
GCGTCGTCGA GCTCGCTCTC GGTGGTCAAT ACCAACTCGT CGCCGCCGTC GTCCTTCGTG
CTCGACATCG CGGTCGACGA TTCCGGCGCC CTGACGGTGT CGGTCGGCGG CGACAGCTCG
CTGTTCACCG TCAGCGGCAA CACCATCATC GGCGCCTCCG GCACGGTGTA TTCCGGCATG
GCCTTCACCT ATTCGGGCTC CAGCTCGGCG TCGATCACCG TGACCTCGAC CTCCGGCATC
GCGGCGCAGA TCAACAACAT CGCCGACCTC GCCTCCGACA CCAGCACGGG GTCGCTGCAG
GATCTGGTCA CCAGCCTGCA ATCGCAGGAC GACCGGATGG AGCAGCAGAT CAACGACATC
AACGAGCGGG CCGAAATCTA CCGCGCGATG CTGGTCAGCC AATACGCCAA ATACCAGAGC
GCGATCTCCA CGGCGGACAC CACGCTCGAC TATCTCTCCG CTCTCCTCAA CGACGAGTAA
 
Protein sequence
MATVTSSTSA SATAALATTT SGTTTTTSVD WDALIEAQVA TKTAAADTIE TSITANEAKI 
SAYQNLQTLL DTLVTSTTSL SKSIVNSLSD STFGARAATI TSSGDVSASS AVSMSISNGA
ATGDHTLTVE QIATAHRVIG TSVADKSADM GLTGVFSLGL AGGTSVDVSI TSGMSMEDIA
DTINAQSDST NVQASIIQIS STEYALTLTA LNDNAEITTS VVSGDDVLTT LGVTDSAGDF
TDVLQEPQPA LFTVDGISLT RDTNDITDVL SGVTFSLLQA TPDGSTINLS IDVDADQIAA
ALEEFVTAYN AVREEVTAQQ TLTSDGTADS SAVLFGDGTM RSIMTQIEQA MNSTVGGLSM
TDLGLSFTDT NTLEFDTSVL SATLTEDLSG VIALLASKTT ASSSSLSVVN TNSSPPSSFV
LDIAVDDSGA LTVSVGGDSS LFTVSGNTII GASGTVYSGM AFTYSGSSSA SITVTSTSGI
AAQINNIADL ASDTSTGSLQ DLVTSLQSQD DRMEQQINDI NERAEIYRAM LVSQYAKYQS
AISTADTTLD YLSALLNDE