Gene Shewmr4_2248 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagShewmr4_2248 
SymbolhemH 
ID4252819 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameShewanella sp. MR-4 
KingdomBacteria 
Replicon accessionNC_008321 
Strand
Start bp2689382 
End bp2690428 
Gene Length1047 bp 
Protein Length348 aa 
Translation table11 
GC content53% 
IMG OID638118873 
Productferrochelatase 
Protein accessionYP_734376 
Protein GI113970583 
COG category[H] Coenzyme transport and metabolism 
COG ID[COG0276] Protoheme ferro-lyase (ferrochelatase) 
TIGRFAM ID[TIGR00109] ferrochelatase 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.000320467 
Plasmid hitchhikingYes 
Plasmid clonabilityhitchhiker 
 

Fosmid Coverage information

Num covering fosmid clones17 
Fosmid unclonability p-value0.140516 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGTTAGCCT GCCGAGGTAT TTGGCTGATA AAAGGTACCA CATTGACTTC TCCCTCTCCT 
GCGTTTGGCG TGTTATTGGT CAATCTCGGC ACGCCCGATG AACCCACTCC CAAAGCGGTT
AAGCGATTCC TCAAACAGTT TTTAAGTGAT CCTCGGGTCG TCGATTTGTC CCCTTGGCTG
TGGCAGCCGA TTTTGCAGGG GATAATCCTT AACACCCGAC CCGCCAAGGT GGCCAAACTT
TATCAGAGCG TGTGGACGGA GCAGGGCTCG CCGCTGATGG TGATAAGCGA GCAGCAGGCG
CAGAAGTTAG CCACGGATCT GAGCGCGACC TTTAATCAAA CCATTCCGGT GGAACTGGGC
ATGAGCTATG GCAATCCTTC GATTGATAGC GGCTTTGCCA AACTTAAGGC CCAAGGCGCC
GAACGTATCG TGGTACTGCC GCTGTATCCG CAATATTCCT GCTCGACCGT CGCCAGTGTG
TTCGATGCGG TGGCGCAGTA TTTTACCCAA GTGCGTGACA TTCCTGAGCT GCGTTTCAGC
AAACAGTATT TTGACCATGA CGCCTATATC GCGGCCTTAG CGCATTCGGT TAAGCGCCAT
TGGAAAACCC ATGGGCAGGC CGATAAGTTG ATTTTATCCT TCCACGGTAT TCCGCTGCGT
TATGCCACCG AAGGCGATCC CTATCCTGAG CAGTGCCGCT CGACGGCTAA GTTATTGGCG
CAGGCGCTGG AGTTAACCGA CGGACAATGG CAGGTGTGTT TCCAATCCCG CTTCGGTAAA
GAAGAGTGGT TAACCCCCTA TGCCGATGAG CTGCTGGCCG ATTTACCCCG CCAAGGCGTA
AAAAGTGTCG ATGTCATTTG CCCAGCCTTT GCCACCGATT GCCTTGAAAC TTTAGAAGAA
ATTTCCATTG GCGGTAAAGA GACTTTCCTG CATGCGGGCG GCGAGGCCTA TCACTTTATT
CCCTGTTTAA ATGATGATGA GCTCCATATC GAGCTGCTCA GGCAACTGGT GCAAGAACAA
GCTCAACCTT GGATATGCGC AGAGTGA
 
Protein sequence
MLACRGIWLI KGTTLTSPSP AFGVLLVNLG TPDEPTPKAV KRFLKQFLSD PRVVDLSPWL 
WQPILQGIIL NTRPAKVAKL YQSVWTEQGS PLMVISEQQA QKLATDLSAT FNQTIPVELG
MSYGNPSIDS GFAKLKAQGA ERIVVLPLYP QYSCSTVASV FDAVAQYFTQ VRDIPELRFS
KQYFDHDAYI AALAHSVKRH WKTHGQADKL ILSFHGIPLR YATEGDPYPE QCRSTAKLLA
QALELTDGQW QVCFQSRFGK EEWLTPYADE LLADLPRQGV KSVDVICPAF ATDCLETLEE
ISIGGKETFL HAGGEAYHFI PCLNDDELHI ELLRQLVQEQ AQPWICAE