Gene Haur_4351 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagHaur_4351 
Symbol 
ID5736211 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameHerpetosiphon aurantiacus ATCC 23779 
KingdomBacteria 
Replicon accessionNC_009972 
Strand
Start bp5559523 
End bp5561259 
Gene Length1737 bp 
Protein Length578 aa 
Translation table11 
GC content50% 
IMG OID641281512 
ProductAMP-dependent synthetase and ligase 
Protein accessionYP_001547111 
Protein GI159900864 
COG category[I] Lipid transport and metabolism
[Q] Secondary metabolites biosynthesis, transport and catabolism 
COG ID[COG0318] Acyl-CoA synthetases (AMP-forming)/AMP-acid ligases II 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones12 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
GTGGATACGC CATGGACAAC ACACTACGAG CCAACGGTCA AGCCCTCATT GCAGTACTCC 
ACTGAGCCGT TGATTTCGTT ATTGGATAGC GCCGTTGCCA ATTATCCCAA TCAGGTGGCT
ACAAACTTTG TACTGAAATA TGTACTTGGT GGTCGGGTGG CAATTGGCGG TTCGCTCACT
TATCGGCAAT TGAAGGAAGC GGTTGATCGG TTTGCCACAG CCTTGCATGG CTTGGGTGTT
CGCAAGGGCG ATCGCTTTGC GATTATGTTG CCCAACACGC CTCAGTTTGT AATTGGTTTT
TTCGCGGCGC TGCGTTTGGG AGCCACGGTT GTCAACATCA ACCCAACCTA CTCGCCACGC
GAACTCAAGC ATCAATTGGC CGACTCTGGC GCTGAGACAA TTTTTGTGCT TAGCCCGTTC
TACCCTAAAG TTCAAGAAAT TTTGGCCGAA ACCCCGCTCA AACGCGTGAT CGTCACCTAC
GTTCACGATG TTTTATCCTT CCCCCAAAGT ATACTTGTGC GGCGCACCCA GCAAAAAGAC
CCCAGTTGGG TTGAGATTAC CCCCGATCGC ACGACCTTGA TGTTTAGCAC CTTGCTGGCT
GAATATCCAC CCGCCCCGCC TAAAGTGGTG ATGGATGGCC ATGATGTGGC GCTGTTTCAA
TATACTGGCG GTACAACTGG CGCTCCCAAG GCAGCGATGC TGACCCACTA CAACATTATG
GCCAACACCA GCCAATTGTT GAATTGGATG AACGACCTTA AAGCTGGCCA AGAGCGTGTC
ATGTGTGCTA TTCCGTTTTT CCATGTCTAT GGGATGACGG TCGGCATGTG CTTTGCGGTG
GCAATTGGCG CAGAAATGGT GATCATTCCT AATCCACGGC CAGTTGATGG TGTGTTGGAA
GCCTTACACA ACGAACGCGC AACGATCTTC CCAGGTGTGC CAGCGCTCTA CATTGGGATT
ATCAATCACA AAGATATTGA TAAATACAAT TTGCGCTCGA TCAAGGCCTG TATCAGTGGC
TCAGCGGCCT TGCCAATGGA AGTTCAAGAG AAATTTGGTG AATTGACTGG CGGGCGTTTG
GTCGAAGGCT ACGGCCTGAC CGAAGCGGCT CCAGTTACTC ACTGTAACCC AGTTTTTGGC
ACGCGCAAAT CTGGCTCAAT TGGCGTACCA ATGCCCGATG TTGAAGTGCA AATTTTGGAT
CTTGAGACCG AAGCACCGTT GCCAATTGGC TCAGAGCGTG AAGGCGAATT GCTGATTCGT
GCGCCGCAAA TTATGAAGGG CTATTGGGGT CGCGATGATG AAACCGCCAA AGTACTGACC
GAAGATGGTT GGCTGCGCAC TGGCGATATT GTCAAAACCG ATAGCGATGG TTATTTCTAT
GTGGTCGATC GCAAGAAAGA CTTAATTATC GCTTCGGGCT ATAATATTGT GCCACGTGAA
GTCGAGGAAG TGCTGTTTAT GCACCCAGCC GTGCTTGAAG CGGCGGTGGC TGGCATTCCT
GATACCTACC GTGGTGAAAC AGTTAAAGCC TTTGTGGTGC TCAAGCCCGA TGCAGCTGCC
ACCGCCAAAG AAATTCGCGA TTTTTGCAAA GAAAATCTCG CACCTTATAA AGTTCCAACT
CAAGTGGAAT TTATTGAGGA ACTGCCCAAA ACCCAAGTTG GCAAAGTGCT ACGGCGGGTA
TTGGTCGAGA TGGAAAAGCA AAAACAAGCT GAGCAAATCC AGCAGGAGGC AAGTTAG
 
Protein sequence
MDTPWTTHYE PTVKPSLQYS TEPLISLLDS AVANYPNQVA TNFVLKYVLG GRVAIGGSLT 
YRQLKEAVDR FATALHGLGV RKGDRFAIML PNTPQFVIGF FAALRLGATV VNINPTYSPR
ELKHQLADSG AETIFVLSPF YPKVQEILAE TPLKRVIVTY VHDVLSFPQS ILVRRTQQKD
PSWVEITPDR TTLMFSTLLA EYPPAPPKVV MDGHDVALFQ YTGGTTGAPK AAMLTHYNIM
ANTSQLLNWM NDLKAGQERV MCAIPFFHVY GMTVGMCFAV AIGAEMVIIP NPRPVDGVLE
ALHNERATIF PGVPALYIGI INHKDIDKYN LRSIKACISG SAALPMEVQE KFGELTGGRL
VEGYGLTEAA PVTHCNPVFG TRKSGSIGVP MPDVEVQILD LETEAPLPIG SEREGELLIR
APQIMKGYWG RDDETAKVLT EDGWLRTGDI VKTDSDGYFY VVDRKKDLII ASGYNIVPRE
VEEVLFMHPA VLEAAVAGIP DTYRGETVKA FVVLKPDAAA TAKEIRDFCK ENLAPYKVPT
QVEFIEELPK TQVGKVLRRV LVEMEKQKQA EQIQQEAS