Gene Htur_4448 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagHtur_4448 
Symbol 
ID8745077 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameHaloterrigena turkmenica DSM 5511 
KingdomArchaea 
Replicon accessionNC_013745 
Strand
Start bp30891 
End bp32705 
Gene Length1815 bp 
Protein Length604 aa 
Translation table11 
GC content62% 
IMG OID646514985 
Productextracellular solute-binding protein family 5 
Protein accessionYP_003405932 
Protein GI284172550 
COG category[E] Amino acid transport and metabolism 
COG ID[COG0747] ABC-type dipeptide transport system, periplasmic component 
TIGRFAM ID[TIGR01409] Tat (twin-arginine translocation) pathway signal sequence 


Plasmid Coverage information

Num covering plasmid clones18 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGGAAGACG GTAATAGCTG TCGGAATGGC GATCACAGTC GAAGAGATGT CCTAAAATAC 
GGCGCTGTGG GTGGCACGGC TCTCGTCGCT GGCTGTCTCG GTGGGGGAAG TAGCACGGAT
CGGTTCCGGG TGTTCGATCC GCAGTCGAGC GGGACGCTCC CGTCGGAGCG CCACTGTAAT
CCGTTCAACC CGACCCAACG CGGGACGTGG CATCCGGGGG CACTCATCTT CGACCGTCCC
GCCATCCACA GTCCGGCGGA AGACGAGGTC TATCCCCTGG TCGCGACCGA CTGGGAGATG
GTCGACGATA CCACGCTGGA GTTCACGTTC AGCGACGAGT GGACCTGGCA CAACGGCGAT
CAACTCGTCG CCGACGACTG GGTTATGCAG CTCCAGATGG CGCTCGCGAT TCTCGAGCAC
CAGGCCGAGG ACGGCGACCG TCCGCACCAG TTCATCGAGT CCGCCGAAGC GCCCGACGAG
CAGACGGTAC AGATCAGCCT CCACGATCCG CTCTCGGAGG CGGTCGCGGT CCAGAACGCT
ATCGCGGACC TCGTCGGCGA CGAGAGCCGC GGCATCTTCA CCAAACACGA CGACGACCAG
TGGAGCGAGT GGCACAGCCA GCTCATGGAG GCCGACGACT CCGAGATGGA ATCGCTCCTC
GAGGAGCTCA CATCGGAGGG ATATCCGTTA CTCGAGGACG CGATCGGTAA CGGCCCGTTC
GAGGTCGCCG ACATCGGTGA CAACGTAATG GTCTTCGAGA AGTACGAGGA CCATCCGAAC
GCTGACAACA TCAACTTCAG CGAGTACTCG GTCCACCTCT ACGAGAACAA CAACCCAACC
CAACCGTACG CTAACGGCGA GGTCGACGCC GCACACACGC AGTTCCCAGT CGAAGACGAC
GTCAAAAGCC AACTCCCCGG GGGACACACG CTCATCAAGG AGAGCTTCTC GACCAACAAG
CTGTTCTCGT TCAACTGCGG ACACGACGTT TCCTACGACA CGTACCTCTC GAACGCGAAC
GTCCGCAAGG CGGTCTGCCA CGTCTTCGAC CGCCAGCAGG TCACCGAGGT CCTCGAAGGC
GTCAACCGGA TGTTCGACTG GCCGTCGTGT CGCGTCCCCG GAAACGTGCT GGACAGCGGT
TCCCACGACG CCGCGGAGTG GATTCAGGAC TTCACCGAGT ACGGCCAGAA CGACACCGAA
CGCGCCGCTG AACTCCTCGA ACGGGAGGGA TTCCAGCGAG ACGACGGTGC GTGGTATACG
CCGGACGGTG ACCGGTTCGA GATCGACGTC CTGGGTGGGA CCAAACGGAA GGACTTCGGC
GTTCTCAAGG ACAATCTGAA CGAGTTCGGT ATCGCGACGA ATCAGGAAGA AGTCGACGAC
GCGACGCTCT CCGAGCGCCG ACAGAACGGC GAGTTCGACA TCGTGCCCGA CGGCTCCTCG
GCCAACGGCG TCCGGGCGAT GTGGGCGCTG GATCTCGTGC CGGGTTGGCT CAGTCAGATC
ACGCATTTCG ATCCCAACGC GGAGATTCCG ATGCCCGTCG GCGACCCCGA GGGCTCGAGC
GGGACGAAGA CGTTCAACGT CGAGGAACAC ATTCGGGAGT GGCAGGTCAC CGACGACGAT
CAGTACCACA AGGAACTGAT GTGGTGGTGG AACCAGACCG TTCCACAGAT GGAAACGATG
TATCAGCCCG ACGCCGGCGC CTACAACGCC GACAACTGGG AACTCGACGC GCCCGACGGC
GTCATCGACG GCACCGAGGA CGCGCTCTAC CTCATCCCGA AGATGGACGA GGCGAGCATG
GAGTACACGG GCTGA
 
Protein sequence
MEDGNSCRNG DHSRRDVLKY GAVGGTALVA GCLGGGSSTD RFRVFDPQSS GTLPSERHCN 
PFNPTQRGTW HPGALIFDRP AIHSPAEDEV YPLVATDWEM VDDTTLEFTF SDEWTWHNGD
QLVADDWVMQ LQMALAILEH QAEDGDRPHQ FIESAEAPDE QTVQISLHDP LSEAVAVQNA
IADLVGDESR GIFTKHDDDQ WSEWHSQLME ADDSEMESLL EELTSEGYPL LEDAIGNGPF
EVADIGDNVM VFEKYEDHPN ADNINFSEYS VHLYENNNPT QPYANGEVDA AHTQFPVEDD
VKSQLPGGHT LIKESFSTNK LFSFNCGHDV SYDTYLSNAN VRKAVCHVFD RQQVTEVLEG
VNRMFDWPSC RVPGNVLDSG SHDAAEWIQD FTEYGQNDTE RAAELLEREG FQRDDGAWYT
PDGDRFEIDV LGGTKRKDFG VLKDNLNEFG IATNQEEVDD ATLSERRQNG EFDIVPDGSS
ANGVRAMWAL DLVPGWLSQI THFDPNAEIP MPVGDPEGSS GTKTFNVEEH IREWQVTDDD
QYHKELMWWW NQTVPQMETM YQPDAGAYNA DNWELDAPDG VIDGTEDALY LIPKMDEASM
EYTG