Gene Acid345_3653 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagAcid345_3653 
Symbol 
ID4072256 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameCandidatus Koribacter versatilis Ellin345 
KingdomBacteria 
Replicon accessionNC_008009 
Strand
Start bp4322405 
End bp4323448 
Gene Length1044 bp 
Protein Length347 aa 
Translation table11 
GC content59% 
IMG OID637985676 
Productaldo/keto reductase 
Protein accessionYP_592728 
Protein GI94970680 
COG category[C] Energy production and conversion 
COG ID[COG0667] Predicted oxidoreductases (related to aryl-alcohol dehydrogenases) 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones14 
Plasmid unclonability p-value0.233218 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones15 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGAGCAGCG ACTTCGACCG GCGGTCTTTT CTAAAAAGCG CGGCAATCAT CGGCGCGGGA 
AGCATAGTTC CGGGGAAAAT GACAGCGATG GCACAGAACA CGAACTCCGA AAAAAGGTCC
GGCGGCGCAA TCCCTAGGCG GAAGCTTGGC AAAACCGGCA TTGAAGTTTC AGCGCTCGGT
ATGGGCGGCT ACCATCTCGG CTCCGCAAAA ACACAGAAGG ATGCGGTCGA GATGGTCGCT
CGCGCCATCG ATGCTGGGAT CACTTTCTTC GACAACGCCT GGGACTATCA CGACGGTCAG
AGCGAGGAAT ATCTCGGCAG CGCGCTGAAA GGGAAGCGCC AGCAGGTCGT CGTCATGACC
AAGGTCTGCA CCCATGGACG CGACAAAAAT GTCGCCATGC AGCAGTTAGA GCAATCGCTG
CGCCGCCTCC AGACCGATCA CCTCGATGTC TGGCAGATCC ACGAAGTTAT TTACGACAAC
GATCCCGACC TCATCTTCCG CCCGGATGGC GCGATCGAAG CGCTCACTCT CGCGAAGCAG
CAAGGGAAAG TTCGCGCCGT TGGTTTTACT GGCCACAAAG ACCCGCGCAT CCACCTCGCG
ATGCTCGCGC ACAATTTCCC CTTCGACACG ATCCAGATGC CGCTGAATTG TCTCGATGCG
ACATTCCGCA GCTTCGAACA GCATGTGCTC CCCGAGGCCC AGAAGCGCGG TATCGCGGTG
CTCGGTATGA AGAGCATGGG CGGCAGCGGC GAGATCGTGA CCCACGGCGC AACGCGTCCC
GACGAAGCTC TGCGCTACGC CATGAGCCTT CCGGTGGCGA CGACGATCAG CGGCATGGAA
TCCATGGAAG TGCTCGAGCA AAACATCGGC ATTGCCTCCG GTTTTCAACC GATGGCCGCG
AATGAAATGC AGTCGCTTCG GGATCAAGTG AAGTATTGGG CCGCCGATGG CCGCTTTGAG
CGCTTCAAAA CCACCAAGAT GTACGACGGC GCCGAAGGCC GCAAACAGCA CGAATATCCG
CCCACCAGCG AATTGCCGGC TTGA
 
Protein sequence
MSSDFDRRSF LKSAAIIGAG SIVPGKMTAM AQNTNSEKRS GGAIPRRKLG KTGIEVSALG 
MGGYHLGSAK TQKDAVEMVA RAIDAGITFF DNAWDYHDGQ SEEYLGSALK GKRQQVVVMT
KVCTHGRDKN VAMQQLEQSL RRLQTDHLDV WQIHEVIYDN DPDLIFRPDG AIEALTLAKQ
QGKVRAVGFT GHKDPRIHLA MLAHNFPFDT IQMPLNCLDA TFRSFEQHVL PEAQKRGIAV
LGMKSMGGSG EIVTHGATRP DEALRYAMSL PVATTISGME SMEVLEQNIG IASGFQPMAA
NEMQSLRDQV KYWAADGRFE RFKTTKMYDG AEGRKQHEYP PTSELPA