Gene RoseRS_3539 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagRoseRS_3539 
Symbol 
ID5210517 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameRoseiflexus sp. RS-1 
KingdomBacteria 
Replicon accessionNC_009523 
Strand
Start bp4433698 
End bp4435122 
Gene Length1425 bp 
Protein Length474 aa 
Translation table11 
GC content62% 
IMG OID640597135 
Productnickel-dependent hydrogenase, large subunit 
Protein accessionYP_001277847 
Protein GI148657642 
COG category[C] Energy production and conversion 
COG ID[COG3259] Coenzyme F420-reducing hydrogenase, alpha subunit 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones19 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones12 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGACCCGTC ACATCCTGAT CGATCCGGTC ACCCGGATCG AAGGACACGC GAAAATCAGC 
ATCCATCTGG ACGACGACGG AAACGTGGCA GAGGCGCGTT TCCATGTGAC TGAGTTTCGC
GGGTTTGAGC GCTTTTGCGA GGGACGACCA TTCTGGGAAA TGCCCGGCAT TACGGCGCGC
ATCTGCGGGA TCTGCCCGGT CAGCCATCTG CTGGCATCGG CGAAGGCAGG TGACGCGATC
CTCTCGGTAG TGATCCCGCC AGCAGCGGAG AAACTGCGCC GCCTGATGAA CCTGGGGCAG
ATCGTGCAAT CGCACGCGCT GAGTTTCTTT CATCTCAGTG CGCCCGATCT GCTGCTCGGT
TTCGACAGCG ATCCCGCCAC GCGCAATGTC TTTGGATTGA TGGCTGCCGA TCCGACGCTG
GCGCGTGCCG GAATACGGCT TCGCCAGCTG GGGCAGGACA TTATTGCTCT GCTCGGCGGC
AGCAAAATCC ATCCGGCGTG GGCTGTGCCG GGCGGTGTCC GCTCTGCGCC GACCGCCGCA
CAACGCGCCG GGATCATTGA GCGCTTGCCC GAAGCGCGCG CCACGGTGCT CGATGCACTA
CGTCGGTTCA AGGCGTTGCT CGACACCCAC GCCGACGAAG TTGCGACATT TGGCAATTTC
CCGTCACTTT TCCTGGGATT GGTAGGACCA AACGGTGAAT GGGAGCACTA CGATGGGCGT
TTGCGTGTGG TTGATTCGGG TGGCGCCATC ATCGCCGATC AGGTTGATCC ATCGCGCTAT
CGCGACATTA TCGCCGAAGC GATCGAGCCG TGGTCATACC TGAAGATGCC GTACTACCGA
CCACGTGGCT ACCCTGGCGG CATGTACCGC GTCGGTCCGC TGGCGCGTCT CAATATCTGC
ACCCGCATCG GCACCCTGCT GGCAGACGCC GAACTGGGGG AGATGCGCCA GCGTGCAGGG
GGTATCGCTA CATCGTCGTT CTACTACCAC TACGCGCGCC TGATCGAGAT TCTGGCGGCG
CTGGAGCGCA TTTCGTTGAT CCTCGACGAT CCTGATCTCG ATTCGCCCCG CCTCCGCGCC
GAAGCGGGAG TCAACCGGTT TGAGGGCGTC GGCGTGAGCG AAGCGCCACG CGGAACCCTC
TTCCACCACT ACACCGTCGA TGCGCACGGC TTGATCCAGC GCGTCAATCT GATTATCGCC
ACGGGACACA ACAATCTGGC GATGAACCGG ACGATTGCCC AAATCGCGCG GCACTTTGTG
CACGGCGATC GGATCGGCGA AGGGGCGTTG AACCGGGTGG AGGCAGGCAT CCGCGCCTAC
GATCCGTGTC TCAGTTGTTC GACGCATGCC GCTGGAACGA TGCCGCTGAC GCTCACACTG
GTCGCCGCTG ATGGTACGGT GCTCGATGAG GTGCGGCGGG GGTGA
 
Protein sequence
MTRHILIDPV TRIEGHAKIS IHLDDDGNVA EARFHVTEFR GFERFCEGRP FWEMPGITAR 
ICGICPVSHL LASAKAGDAI LSVVIPPAAE KLRRLMNLGQ IVQSHALSFF HLSAPDLLLG
FDSDPATRNV FGLMAADPTL ARAGIRLRQL GQDIIALLGG SKIHPAWAVP GGVRSAPTAA
QRAGIIERLP EARATVLDAL RRFKALLDTH ADEVATFGNF PSLFLGLVGP NGEWEHYDGR
LRVVDSGGAI IADQVDPSRY RDIIAEAIEP WSYLKMPYYR PRGYPGGMYR VGPLARLNIC
TRIGTLLADA ELGEMRQRAG GIATSSFYYH YARLIEILAA LERISLILDD PDLDSPRLRA
EAGVNRFEGV GVSEAPRGTL FHHYTVDAHG LIQRVNLIIA TGHNNLAMNR TIAQIARHFV
HGDRIGEGAL NRVEAGIRAY DPCLSCSTHA AGTMPLTLTL VAADGTVLDE VRRG