Gene Haur_0259 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagHaur_0259 
Symbol 
ID5732154 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameHerpetosiphon aurantiacus ATCC 23779 
KingdomBacteria 
Replicon accessionNC_009972 
Strand
Start bp302771 
End bp303808 
Gene Length1038 bp 
Protein Length345 aa 
Translation table11 
GC content49% 
IMG OID641277383 
Productsulfate ABC transporter, periplasmic sulfate-binding protein 
Protein accessionYP_001543039 
Protein GI159896792 
COG category[P] Inorganic ion transport and metabolism 
COG ID[COG1613] ABC-type sulfate transport system, periplasmic component 
TIGRFAM ID[TIGR00971] sulfate/thiosulfate-binding protein 


Plasmid Coverage information

Num covering plasmid clones14 
Plasmid unclonability p-value0.904874 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGAAACGCT TGCATTTCAG TTTATTAAGC TTGTTATTGG TCGGATTATT GGCTGCTTGT 
GGCGAGGCCA GCCAAACTAC CACCAGCAAT GCCACGGTTA CCACCATCAC CTTGGGCGCG
TACACCACGC CACGCGAAGC CTATGCCAAG TTGATTCCGC TGTTTCAAGC CAAATGGAAG
GCCGATACTG GCGGTGAAGT CAAATTTGAA GAATCATATC AAGGTTCAGG TGCGCAATCA
CGGGCAATCG TTGAAGGCTT CGAGGCTGAT ATCGCGGCGC TTTCGCTCGA AGCTGATATT
AATCGGATCA CCGATGCAGG CCTGATCACT CACGATTGGA AAGCTGGCAC GCATTCGGGC
ATGGTTAGTA CCTCAATTGT GGTGTTTGCG GTACGCGAAG GCAATCCCAA AGGTATTCAA
GATTGGGCTG ATTTAGCCAA GCCTGGAGTA CAAATTTTGA CTCCCGATCC ACGCACCAGC
GGCGGCGCAC AATGGAATAT TTTAGCGTTG TATGGGGCCG CCAAACGTGG TCAGATTACA
GGCGTACCCG CCAACGACGA AGCAGCAGCC CAAGCCTTTT TAGCGAGTGT GCTTAAAAAT
GTCGTGGTGT TTGATAAGGG TGCGCGTGAA AGTATCACCA ACTTTGAAAA AGGCGTTGGC
GATGTGGCAA TCACCTATGA AAATGAAATT TTGGTTGGCC AAAAAGGCGG CCAAACCTAT
CAAATGGTCA TCCCAACTTC CACGATTTTG ATCGAAAACC CAATTGCCCT AATCGATAAA
TCAGTTGAAA AACATGGCAA TCGTCAAGCA GTCGAAGCTT TTATTAACTT CTTGCATAGC
CGTGAAGCCC AAGAAGTCTT TGCTGAATTT GGCTTACGCT CGGTCGATGC CGATGTTGCC
AAAGCCACTG CTGAGCGCTA TCCCGCCATC AACGATTTGT TTACAATCAA CGAATTTGGT
GGTTGGAGCA AAGCCACGCC TGAATACTTT GGCGATGACG GTGTCTATGC CAAAGTGCTA
GCGCAGGTAC AACAATGA
 
Protein sequence
MKRLHFSLLS LLLVGLLAAC GEASQTTTSN ATVTTITLGA YTTPREAYAK LIPLFQAKWK 
ADTGGEVKFE ESYQGSGAQS RAIVEGFEAD IAALSLEADI NRITDAGLIT HDWKAGTHSG
MVSTSIVVFA VREGNPKGIQ DWADLAKPGV QILTPDPRTS GGAQWNILAL YGAAKRGQIT
GVPANDEAAA QAFLASVLKN VVVFDKGARE SITNFEKGVG DVAITYENEI LVGQKGGQTY
QMVIPTSTIL IENPIALIDK SVEKHGNRQA VEAFINFLHS REAQEVFAEF GLRSVDADVA
KATAERYPAI NDLFTINEFG GWSKATPEYF GDDGVYAKVL AQVQQ