Gene RPB_1047 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagRPB_1047 
Symbol 
ID3908899 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameRhodopseudomonas palustris HaA2 
KingdomBacteria 
Replicon accessionNC_007778 
Strand
Start bp1203970 
End bp1204890 
Gene Length921 bp 
Protein Length306 aa 
Translation table11 
GC content66% 
IMG OID637882940 
Productsulfate ABC transporter, permease protein CysW 
Protein accessionYP_484668 
Protein GI86748172 
COG category[P] Inorganic ion transport and metabolism 
COG ID[COG4208] ABC-type sulfate transport system, permease component 
TIGRFAM ID[TIGR00969] sulfate ABC transporter, permease protein
[TIGR02140] sulfate ABC transporter, permease protein CysW 


Plasmid Coverage information

Num covering plasmid clones20 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones27 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGACGCTGG TCGCGACCAC ATCCGTGAGG CGCTCCCCGC GCAAGCCTGC CGTCGCGGCC 
GGCGCACGTG CCGGGCAGCC GGCGCGGCGC CCGGCGCATG GCGAGCCGGC CTGGGTGCGC
CTGCTGATCA TCGGCTTCGC CGTGAGTTTT CTCACCGTCT TCGTGGTGCT GCCGCTGATC
CTGGTGTTCT CGGAGGCGCT GTCGAAGGGC GTCTCGTTCT ATCTCGACGC GCTGGCGGGA
GACGAAGCGC TGGCGGCGAT CCGGCTGACG CTGGTCGCGG CGGCGATCTC GGTCGGGCTC
AATCTGGTGT TCGGCGTGAT CGCCGCCTGG GCGATCGCGA AGTTCGAGTT TCGCGGCAAG
ACGCTGCTGA TCACGCTGAT CGATCTGCCG TTCTCGGTCA GCCCGGTGAT CTCCGGCCTG
GTATTCGTGC TGCTGTTCGG CGCGCAGGGC TTTGTCGGCC CGTGGCTGAT GGCCCACGAC
GTGCGAATCC TGTTCGCGCT GCCGGCGATC GTGCTGGCGA CCACCTTCGT GACCTTCCCG
TTCGTCGCGC GCGAACTGAT CCCGCTGATG CAGGAGCAGG GCCAGCACGA GGAAGAAGCC
GCGATCTCGC TCGGCGCCAG CGGCTGGAAA ACCTTCTGGC GGGTGACGCT GCCGAACATC
AAATGGGGCC TGCTGTACGG CGTGCTGCTG TGCAATGCGC GGGCGATGGG CGAGTTCGGC
GCGGTGTCGG TGGTGTCGGG TCACATCCGC GGCGAGACCA ACACCATGCC GCTGCTGGTC
GAAATTCTCT ACAACGAGTA TCAGATGGTC GCCGCCTTCG CGATCGCCTC GCTGCTGGCG
CTGCTGGCGC TGGTGACGCT GATCGTCAAG ACCATCTTGG AAGGCCGTAT CGAGGAAGGG
CTGCACACCG ATGACCATTG A
 
Protein sequence
MTLVATTSVR RSPRKPAVAA GARAGQPARR PAHGEPAWVR LLIIGFAVSF LTVFVVLPLI 
LVFSEALSKG VSFYLDALAG DEALAAIRLT LVAAAISVGL NLVFGVIAAW AIAKFEFRGK
TLLITLIDLP FSVSPVISGL VFVLLFGAQG FVGPWLMAHD VRILFALPAI VLATTFVTFP
FVARELIPLM QEQGQHEEEA AISLGASGWK TFWRVTLPNI KWGLLYGVLL CNARAMGEFG
AVSVVSGHIR GETNTMPLLV EILYNEYQMV AAFAIASLLA LLALVTLIVK TILEGRIEEG
LHTDDH