Gene Rsph17025_1751 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagRsph17025_1751 
Symbol 
ID5083657 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameRhodobacter sphaeroides ATCC 17025 
KingdomBacteria 
Replicon accessionNC_009428 
Strand
Start bp1787236 
End bp1788498 
Gene Length1263 bp 
Protein Length420 aa 
Translation table11 
GC content66% 
IMG OID640483311 
ProductHK97 family phage portal protein 
Protein accessionYP_001167949 
Protein GI146277790 
COG category[S] Function unknown 
COG ID[COG4695] Phage-related protein 
TIGRFAM ID[TIGR01537] phage portal protein, HK97 family 


Plasmid Coverage information

Num covering plasmid clones14 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones18 
Fosmid unclonability p-value0.518591 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGGGCCTTT TCGACTTCTT CCGGAGCGAG CCGCAGGTGG CCCGCGTGGA GCCGCCTGTG 
GTGGCGCAGG GCTCCGGCGA TGTGCAGAGC CCCGGCCAAT GGCGCGGCTT CGTCACCGGC
GGCGTCTCGC GCTCCGGGGT GCGGGTGAAC GAGACGACGG CGCTTTCGAT CCCCGCCACG
CTTCAGGCGA TCCGGGTCCT CTCCGGTGTG TTCGCCATGA CGCCGCTGCA CTATTTCCGC
CGCACCGGTG ATGGCCGCGA GCGCGTGTCG GATGACATCG CAGCCCTCCT TCACGACCGG
CCGAACAGCC ATCAGACCGC GTTCGCCTTC CGCGAACTGC TCAAGATGGA CCTGCTGCTG
TCGGGGAACT TCTACGCCTA TGTCAGCCGC GACTTCGCCG GCCGTCCGAA GGCGCTGACG
CGCCTCAAGC CCGGCAGCGT CCTGATTGCG GAGTACTTCG ATCGCTCGGA GGGGGTCACG
CTCTTTTATG ATGCAACCTT GCCGGACGGG TCGCGGGAGA GGTTTCCCGC CCGGGACATC
TGGCACATTG CAGGCATGAG CCGTGATGGG CTGTCCGGAC TGAACCCGAT CCAGTTCGCG
CGCGACGCCA TCGGCGGGGC CATCGCCACG GCTGACCATG CCGCGAAGTT CTGGGGGAAC
GGGGGGCGTC CAAGCACCCT GCTGAAGACC AAGCACAAGG TGGACCCGAT CGCGCGAAAG
CAGATCAAGT CCGACTGGAA GGCGATCTAC GGCGGACCGT TCGGCGACGA CATTGCCGTC
CTCGACCAGG AGTTGGAGGC CCAGTTCCTC AGCCACGACA ACAAGGCGTC GCAGTACCTT
GAGACGCGCG GCTTTCAGGT CATGGACCTG GCGCGCCTCT GGGGCGTGCC GCCGCATCTG
ATCTTCGACC TGTCGAGGGC CACCTTCTCG AATATCGAGC AGCAGAGCCT CGAGTTCATC
GTGTTCCACC TCGGCCCGCA CTACGAGCGG GTGAGCCAGT CGGCCACGCG CCAGTTCGCC
GCGGATGCCC ATTATTTCGA ACATGTCACC GACGCTCTGG TGAAGGGCGA TGTGAAGAGC
CGCATGGAGG CCTACTGGCT CCAGCGGCAG ATGGGCATGG TCAACGCCAA CGAGCTGCGT
CGGCGCGACA ACCTCTCGCC GATCTCTGGC GATGCCGGCG AGGAATACTG GCGTCCCGCC
GCCATGACGC TGGCGGGCAC GCCGCCAGAG CAACCCGCGC AGCGGGCCTC GGCAGAACCC
TGA
 
Protein sequence
MGLFDFFRSE PQVARVEPPV VAQGSGDVQS PGQWRGFVTG GVSRSGVRVN ETTALSIPAT 
LQAIRVLSGV FAMTPLHYFR RTGDGRERVS DDIAALLHDR PNSHQTAFAF RELLKMDLLL
SGNFYAYVSR DFAGRPKALT RLKPGSVLIA EYFDRSEGVT LFYDATLPDG SRERFPARDI
WHIAGMSRDG LSGLNPIQFA RDAIGGAIAT ADHAAKFWGN GGRPSTLLKT KHKVDPIARK
QIKSDWKAIY GGPFGDDIAV LDQELEAQFL SHDNKASQYL ETRGFQVMDL ARLWGVPPHL
IFDLSRATFS NIEQQSLEFI VFHLGPHYER VSQSATRQFA ADAHYFEHVT DALVKGDVKS
RMEAYWLQRQ MGMVNANELR RRDNLSPISG DAGEEYWRPA AMTLAGTPPE QPAQRASAEP