Gene Haur_1464 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagHaur_1464 
Symbol 
ID5733349 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameHerpetosiphon aurantiacus ATCC 23779 
KingdomBacteria 
Replicon accessionNC_009972 
Strand
Start bp1706886 
End bp1707974 
Gene Length1089 bp 
Protein Length362 aa 
Translation table11 
GC content52% 
IMG OID641278602 
ProductnifR3 family TIM-barrel protein 
Protein accessionYP_001544236 
Protein GI159897989 
COG category[J] Translation, ribosomal structure and biogenesis 
COG ID[COG0042] tRNA-dihydrouridine synthase 
TIGRFAM ID[TIGR00737] putative TIM-barrel protein, nifR3 family 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.00679382 
Plasmid hitchhikingYes 
Plasmid clonabilityhitchhiker 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGCAGCTTA TGCCTGATAC TCTGACCCAA GCATTGCCCG AATCATTTCA ACTTGGCTCA 
ATGCGCATTT TTCCAAACAT GGTGCTGGCC CCGATGGCTG GCGTAACCGA TTCAGTCTTT
CGGCGGCTGG TGTTGTCGTT GGGTGGCTGT GGCTTGGTTG TTTCTGAGAT GACCAATGCT
GCCAGTGTTT CGCCCAAGGC CATGAAACGC CATCGCTTGC TGGATTATCT GCCCGAAGAA
CGGCCAATTT CGATTCAACT TTCGGGCAAC GACCCCGATT TGGTGGCAAC CGCCGCCCGC
TTTGTTGAAG AACTTGGCCC CGATGTAATT GACATAAACT GTGGCTGCCC TTCGCCCAAA
GTCACTGGTG GCGGCCATGG TTCGGCCTTG CTCAAAGATT TGCCCAAAAT GCAGCAAATG
CTCAAAGCCG TGTATGCAGC GATCAACATT CCGTTTACGC TCAAATTTCG GGCTGGCTGG
GATGAGCAAT CGCTAAATTA TATTGATACC GCTAAAATTG CTGAAGATGC TGGTTGTGCC
GCGATTACCT TGCATCCACG TACCAAAGTG CAAGGCTACA GCGGCGATGC CGATTGGTCG
CGGGTTGCTG AAGTTGTGCA AGCAGTCTCG ATTCCGGTGA TTGGCTCAGG CGATGTGCGC
ACACCAGCCG ATGCTTTGGC GCGTTTGGAG CAAACTGGGG TTGCAGCAGT GATGATTGGT
CGAGCTGCCA TGGCCAATCC ATGGATTTTT CGCCAAATTG CTCAATTGCG GGCAGGCGAA
CCCATATTTG TGCCAACACC AAGCGATAAA CGTGATTTGC TGGTGCGCTA TGTTGATATG
TGTGCTGAAA CTATGGTTGA GCGCCAAGCC CTTGGCAAGC TAAAACAACT GATTGGTCAA
TTCAGCATTG GTTTGTATGC TAGCAACCAA TTGCGCCGCG ATGTACAACG TGCCAACGAA
ATTGAAGCCG CCAAAGCGAT TATTGCCAAC TTTTTTGAGC CATTTATCAG CGGCGCGGTC
GAGGCGGTCG AAGTGCCCGA CGAAATCGCC GTGATCAAAG AAGGTTGCGA GAACGGCGCG
AATAATTAG
 
Protein sequence
MQLMPDTLTQ ALPESFQLGS MRIFPNMVLA PMAGVTDSVF RRLVLSLGGC GLVVSEMTNA 
ASVSPKAMKR HRLLDYLPEE RPISIQLSGN DPDLVATAAR FVEELGPDVI DINCGCPSPK
VTGGGHGSAL LKDLPKMQQM LKAVYAAINI PFTLKFRAGW DEQSLNYIDT AKIAEDAGCA
AITLHPRTKV QGYSGDADWS RVAEVVQAVS IPVIGSGDVR TPADALARLE QTGVAAVMIG
RAAMANPWIF RQIAQLRAGE PIFVPTPSDK RDLLVRYVDM CAETMVERQA LGKLKQLIGQ
FSIGLYASNQ LRRDVQRANE IEAAKAIIAN FFEPFISGAV EAVEVPDEIA VIKEGCENGA
NN