Gene Rsph17029_1130 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagRsph17029_1130 
Symbol 
ID4895275 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameRhodobacter sphaeroides ATCC 17029 
KingdomBacteria 
Replicon accessionNC_009049 
Strand
Start bp1175113 
End bp1176285 
Gene Length1173 bp 
Protein Length390 aa 
Translation table11 
GC content70% 
IMG OID640111716 
ProductHK97 family phage portal protein 
Protein accessionYP_001043012 
Protein GI126461898 
COG category[S] Function unknown 
COG ID[COG4695] Phage-related protein 
TIGRFAM ID[TIGR01537] phage portal protein, HK97 family 


Plasmid Coverage information

Num covering plasmid clones14 
Plasmid unclonability p-value0.868122 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones18 
Fosmid unclonability p-value0.923959 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGCTGTTCG ACTTCCTGAG GAAACCGGCG CGCGCTCCGG CGCCCGAGCG CAAGGCCTCG 
GCCACGGGGC CGGTGGTGGG CTGGAGCACG GGGCGCGTGG CCTGGAGCGC GCGGGACATG
GTGTCGCTGA CCCGGAACGG GTTTCTGGGC AATCCGATCG CCTTCCGGTC GGTCAAGCTG
ATCTCGGAGG CGGCGGCCGC GCTGCCTCTG GTTCTGCAGG ATGCGGGGCG GCGCTACGAG
AGCCACCCGA TGCTGGATCT GATCGCGCGG CCGAACCCGT TGCAGGGGCG GGCCGAGCTG
CTCGAGGCGC TCTATGCGCA GCTTCTGCTG ACGGGGAATG CCTATCTCGA GGCGGTGGCC
GGATCGGCGC GGCTGCCCGG AGAGCTGCAT CTCCTGCGCT CGGACCGGAT GAGCCTCGTG
CCGGGGCCGG ACGGGTGGCC CGTGGCCTAC GATTATGCGG TGGGCGGGCG GCGCATCCGC
TTCGACATGA CGGCGGGCCT GCCGATCTGC CATATTCGCA CCTTCCATCC GCAGGACGAT
CATTACGGCT TCTCGCCGTT GCAGGCGGCG GCGGTGGCGC TCGACGTGCA TGTGGCGGCT
TCGGCCTGGT CGAAGGCCTT GCTCGACAAT GCCGCCCGGC CCTCGGGGGC CATCGTCTAT
CGCGGCTCGG ACGGGCAGGG AAGCCTGTCC TCGGATCAGT ATGACCGGCT GGTGGGCGAG
ATCGAGGCCA ACCATCAGGG TGCGCGCAAT GCGGGGCGGC CGATGCTGCT GGAGGGCGGG
CTCGACTGGA AGCCGATGGG CTTCTCGCCC TCCGACATGG AGTTCCACAC CACCAAGGAG
GCTGCGGCGC GCGAGATCGC CATCGCCTTC GGCGTGCCGC CGATGCTGCT CGGCATCCCC
GGCGAGGCGA CCTACGCCAA TTATCAGGAG GCGCACCGGG CCTTCTACCG GCTGACGGTG
CTGCCGCTGG CGGCGAAGGT CACGGCCACG TTGTCGCACT GGCTCGGCAG TTTCAGCGGC
GAGGCGGTGG AGCTGCGGCC CGACCTCGAT CAGGTGCCGG CGCTGGCGGC GGAACGGGAT
CAGCAGTGGG CGCGCGTCGC CGCGGCGGAT TTCCTGACGG AGGCCGAGAA GCGGACGCTC
CTCGGCCTGC CGAGGATCGC GGAGGAGGAG TGA
 
Protein sequence
MLFDFLRKPA RAPAPERKAS ATGPVVGWST GRVAWSARDM VSLTRNGFLG NPIAFRSVKL 
ISEAAAALPL VLQDAGRRYE SHPMLDLIAR PNPLQGRAEL LEALYAQLLL TGNAYLEAVA
GSARLPGELH LLRSDRMSLV PGPDGWPVAY DYAVGGRRIR FDMTAGLPIC HIRTFHPQDD
HYGFSPLQAA AVALDVHVAA SAWSKALLDN AARPSGAIVY RGSDGQGSLS SDQYDRLVGE
IEANHQGARN AGRPMLLEGG LDWKPMGFSP SDMEFHTTKE AAAREIAIAF GVPPMLLGIP
GEATYANYQE AHRAFYRLTV LPLAAKVTAT LSHWLGSFSG EAVELRPDLD QVPALAAERD
QQWARVAAAD FLTEAEKRTL LGLPRIAEEE