Gene Dret_1447 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagDret_1447 
Symbol 
ID8419276 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameDesulfohalobium retbaense DSM 5692 
KingdomBacteria 
Replicon accessionNC_013223 
Strand
Start bp1678266 
End bp1679372 
Gene Length1107 bp 
Protein Length368 aa 
Translation table11 
GC content57% 
IMG OID645038022 
Productaminotransferase class I and II 
Protein accessionYP_003198312 
Protein GI258405570 
COG category[E] Amino acid transport and metabolism 
COG ID[COG0436] Aspartate/tyrosine/aromatic aminotransferase 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones14 
Plasmid unclonability p-value0.0268591 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones18 
Fosmid unclonability p-value0.142096 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGCGATTGC AGCCTTTCAA ATTGGAACGG TATTTTGCCA AATACGAATT CAATGTCCGC 
CATCTATTGA GTTCTTCGGA TTGCGAGTCC ATGACCGTGG CGGACTTGCT GGACCTGGAG
CCCGGCGCTG CAGAGCGTTT TCACAATGTC TGGCTCGGCT ATACCGAATC CGAGGGCAGT
CCAACTCTGC GCGAGACGAT CGCTTCCATG TACAATGCGC AGCAGGCCGA CGACATTCTG
GTCCACAGCG GTGCGGAGGA AGCGATTTTT TTGTTCATGA ACGCGGTCCT GGAAGCCGGG
GACCACGTCG TTGTCCACTG GCCGTGCTAC CAATCCCTGA CTGAAGTGCC GCGGTCCATC
GGCTGCGAGG TCGATCTCTG GAAAGCGCGG GAAGAGGCGC AGTGGGGCTT GGATCTCGAG
GAATTGGACG AACTGCTCAA GCCGAACACC AAAGCCATTA TCGTCAATCT TCCGCATAAT
CCCACAGGAT ATCTCATGGA GCCCGAAACG TTTTCGCGGC TTTGCCAGTT GGCTGAAAAC
CGGGATATCC TCCTGTTTTG TGATGAGGTC TATCGCGAAT CGGAATACGA TGTCTCGCGC
CGTCTGCCCG CTGTCTGTGA CTGCTGCCAG ACCGGTGTTT CGCTTGGCGT GACCTCCAAG
ACCTACGGGC TGCCCGGGTT GCGGATCGGC TGGCTGGCCA CACGCCGTCG GGATGTCCTG
GCCGCAGTGG CCCAGTTGAA AGACTATACG ACGATCTGCA ACAGTGCGTC GAGCGAATTT
TTGGCTGAGT TGGCCCTGCG CCACCGGGAA CACCTTGCCG AGCGCAGTGT GCGCTTGATA
CAAACAAACC TTGCTTTGCT GGACGGGTTT TTTGCCCGGC ATGCTGAGCG ATTCGAATGG
CGACGTCCCC ATGCCGGTCC AATCGCGTTC CCGCGTTTGC GTGATGAAGA CGCGGACGAT
TTTTGCCACC AAGCGGTGGA GCAGGCGAGT GTCCTTTTGT TGCCGGGATC GCTTTACGAG
TATCCCGGCG GTGCATTTCG CATTGGCTTT GGCCGGGCCA GTCTCCCTCA GGCCTTGGAA
GCCCTTGAAA ATTTTCTCCA GCGGTAG
 
Protein sequence
MRLQPFKLER YFAKYEFNVR HLLSSSDCES MTVADLLDLE PGAAERFHNV WLGYTESEGS 
PTLRETIASM YNAQQADDIL VHSGAEEAIF LFMNAVLEAG DHVVVHWPCY QSLTEVPRSI
GCEVDLWKAR EEAQWGLDLE ELDELLKPNT KAIIVNLPHN PTGYLMEPET FSRLCQLAEN
RDILLFCDEV YRESEYDVSR RLPAVCDCCQ TGVSLGVTSK TYGLPGLRIG WLATRRRDVL
AAVAQLKDYT TICNSASSEF LAELALRHRE HLAERSVRLI QTNLALLDGF FARHAERFEW
RRPHAGPIAF PRLRDEDADD FCHQAVEQAS VLLLPGSLYE YPGGAFRIGF GRASLPQALE
ALENFLQR