Gene Dret_1874 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagDret_1874 
Symbol 
ID8419717 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameDesulfohalobium retbaense DSM 5692 
KingdomBacteria 
Replicon accessionNC_013223 
Strand
Start bp2147054 
End bp2147932 
Gene Length879 bp 
Protein Length292 aa 
Translation table11 
GC content59% 
IMG OID645038460 
Productdihydrodipicolinate synthase 
Protein accessionYP_003198736 
Protein GI258405994 
COG category[E] Amino acid transport and metabolism
[M] Cell wall/membrane/envelope biogenesis 
COG ID[COG0329] Dihydrodipicolinate synthase/N-acetylneuraminate lyase 
TIGRFAM ID[TIGR00683] N-acetylneuraminate lyase
[TIGR00674] dihydrodipicolinate synthase 


Plasmid Coverage information

Num covering plasmid clones33 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones13 
Fosmid unclonability p-value0.013061 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGCAATTTC AAGGTGCCTT TACCGCCCTT GTCACCCCCT TCAAAAACGA TCAGGTCGAT 
GAAGACGCCT ATCGCAAACT GGTGGAGTGG CAAATAGAGC AGGGTATTGA TGGACTGGTC
CCTTGTGGCA CAACCGGCGA ATCCGCCACG TTATCCCACG AAGAACATAA ACAGGTGATC
AAAATCTGCG TCGACCAGGC CAAGGGACGC GTGCCGGTGT TGGCCGGCGC CGGCTCGAAC
AACACCCGGG AAGCCATCGA CCTGACCCGC TACGCCAAGG AGGCCAAGGC CGATGGCGCC
TTGCTGATCA CTCCCTACTA CAACAAGCCC ACGCCCAATG GGTTGGTGGC CCACTTCAAG
GCCATTGCCC AGGAAGTCTC CATGCCGTTC GTGGTCTACA ATGTTCCCGG CCGCACCGGT
TTGAATGTCC AGGCCCAGAC TATGGCCCGG CTGTTTCATG AAATCCCGGA AGTGGTTGGG
GTCAAAGAGG CCTCCGGGGA TCTCAAACAG ATCGCCGAGG TCGTGGAGGC CTGTGGACCG
GACTTCACCG TCCTCTCCGG TGAGGATTTC ACTGTCTTGC CCTTGCTAGC GGTCGGCGGA
CACGGCGTGA TCTCCGTAAC GTCCAACATC GCGCCCAAAA TGATGTCCGA TATGTGCCGG
GCCTTTCGCG CCGGCGAGCA GAACAAGGCC CTTGGCCTGA GCCTGGAATT GCTGCCGCTC
TGCCGGGGCA TGTTCCTGGA AACCAATCCC ATCCCGGTCA AGACCGCGCT CTCCATGATG
GGCATGATGG ACCTCGAACT CCGCCTGCCC CTGGTCCCGT TGACCCCTGA AAATGAGGCC
GCTCTCAAGG CGCTTCTCAC AGACAAGGGA CTCATTTAG
 
Protein sequence
MQFQGAFTAL VTPFKNDQVD EDAYRKLVEW QIEQGIDGLV PCGTTGESAT LSHEEHKQVI 
KICVDQAKGR VPVLAGAGSN NTREAIDLTR YAKEAKADGA LLITPYYNKP TPNGLVAHFK
AIAQEVSMPF VVYNVPGRTG LNVQAQTMAR LFHEIPEVVG VKEASGDLKQ IAEVVEACGP
DFTVLSGEDF TVLPLLAVGG HGVISVTSNI APKMMSDMCR AFRAGEQNKA LGLSLELLPL
CRGMFLETNP IPVKTALSMM GMMDLELRLP LVPLTPENEA ALKALLTDKG LI