Gene Tpen_0452 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagTpen_0452 
Symbol 
ID4601855 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameThermofilum pendens Hrk 5 
KingdomArchaea 
Replicon accessionNC_008698 
Strand
Start bp411602 
End bp412612 
Gene Length1011 bp 
Protein Length336 aa 
Translation table11 
GC content58% 
IMG OID639773219 
Productputative DNA-binding/iron metalloprotein/AP endonuclease 
Protein accessionYP_919864 
Protein GI119719369 
COG category[O] Posttranslational modification, protein turnover, chaperones 
COG ID[COG0533] Metal-dependent proteases with possible chaperone activity 
TIGRFAM ID[TIGR00329] metallohydrolase, glycoprotease/Kae1 family 


Plasmid Coverage information

Num covering plasmid clones27 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
GTGCTGTTCG GCATGCGCGC GTTAAAGGTG CTAGGCATAG AGTCTACGGC TCACACGTTT 
GGAGTCGGTA TAGCTACTTC TTCTGGAGAT ATTCTGGTCA ACGTTAATCA CACGTATGTT
CCGCGACATG GAGGCATAAA GCCGACGGAG GCCGCCGAGC ACCATAGCAG GGTTGCTCCC
AAAGTCCTCT CCGAGGCGCT CCAGAAAGCG GGTATCAGTG TTGAAGAGGT GGACGCAGTC
GCGGTCGCGC TGGGCCCGGG TATGGGTCCC TGCCTAAGGG TCGGGGCTAC GCTTGCAAGG
TACCTGGCCT TAAAGTTTGG TAAGCCGCTA GTACCGGTTA ACCACGCAAT AGCCCACTTA
GAGATTTCTA GGCTGACTAC GGGGCTGGAG GACCCCGTGT TCGTCTACGT TGCCGGTGGA
AACACGATGG TCACGACTTT CAACGAGGGT AGATACCGGG TATTCGGCGA GACTCTAGAT
ATACCGCTCG GAAACTGCCT CGACACGTTT GCCAGAGAAG TGGGGCTGGG GTTTCCCGGG
GTTCCGCGAG TAGAGGAGCT GGCGCTTAAA GGGCGGGAGT ACATACCCTT ACCGTACACG
GTCAAGGGGC AGGACGTATC TTACTCGGGG TTGCTCACCC ACGCTCTCTC CCTGTACAGA
TCCGGAAGAG CACGGTTAGA GGACGTCTGC TACAGCCTCG TCGAAACCGC CTATTCGATG
CTGGTCGAAG TCGCGGAGAG AGCCTTAGCG CACACCGGTA AGAGCCAGCT CGTCCTCACG
GGCGGCGTTG CCAGGAGCAG GATACTCCTG GAGAAGCTAA GGAGAATGGT CGAGGATAGA
GGCGGAGTGC TCGGGGTTGT CCCGCCTGAG TACGCAGGAG ACAACGGCGC CATGATAGCG
TATACAGGCG CGCTGGCTTT TTCCCACGGT GTCAGAGTGC CGGTTGAGGA AAGCCGTATA
CAGCCCTACT GGAGGGTGGA CGAGGTCGTC ATTCCGTGGC GATCCAGGTG A
 
Protein sequence
MLFGMRALKV LGIESTAHTF GVGIATSSGD ILVNVNHTYV PRHGGIKPTE AAEHHSRVAP 
KVLSEALQKA GISVEEVDAV AVALGPGMGP CLRVGATLAR YLALKFGKPL VPVNHAIAHL
EISRLTTGLE DPVFVYVAGG NTMVTTFNEG RYRVFGETLD IPLGNCLDTF AREVGLGFPG
VPRVEELALK GREYIPLPYT VKGQDVSYSG LLTHALSLYR SGRARLEDVC YSLVETAYSM
LVEVAERALA HTGKSQLVLT GGVARSRILL EKLRRMVEDR GGVLGVVPPE YAGDNGAMIA
YTGALAFSHG VRVPVEESRI QPYWRVDEVV IPWRSR