Gene AnaeK_4105 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagAnaeK_4105 
Symbol 
ID6785575 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameAnaeromyxobacter sp. K 
KingdomBacteria 
Replicon accessionNC_011145 
Strand
Start bp4636853 
End bp4638007 
Gene Length1155 bp 
Protein Length384 aa 
Translation table11 
GC content66% 
IMG OID642765573 
Producttransposase IS4 family protein 
Protein accessionYP_002136438 
Protein GI197124487 
COG category 
COG ID 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones21 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
GTGAGCGACC AACCTCTCGA CCGCGACCGG ATCACGAAGC TCGTCTCCAC CATCTTCGCC 
GAAGACCTGC ACGCCAAGCG TGTGGCGTCG CTGGCGGGAG CCGCGGTCGG TGTCCTGGAG
GGCGCCGCGC TGGGCATCCA CGCCATCGGA AATTCGCTCG CAGTTGCCGA GGGGCTCAAG
TCGAAGCACG CGGTCAAGCA GGTGGACCGG ATGCTGAGCA ACGAGGGCAT CCCGGTGTGG
AGGCTCTTCG GGAGTTGGGT TCCCTGCGTC GTCGGCGACC GGCTGGAGAT CGTCGTGGCG
CTCGACTGGA CCGACTTCGA CGAGGACGAC CAATCGACCA TCGCGCTGTC GATGATCACC
AGTCACGGTC GGGCCACGCC GCTGCTCTGG AAGACGGTGA TGAAGTCGGA GCTGAAGGGA
TGGCGGAACG AGCACGAGGA CGTGCTCCTC GAGCGATTTC GCGAGGTGCT GCCCGAGGGC
GTGAAGGTCA CTGTCCTCGC AGACCGCGGC TTCGGCGACC AAGCTCTCTA CGAACTGCTC
AAGGACCAGC TCGGCTTCGG CTTCATCGTG CGCTTCCGTG GCGTGGTGAA GGTGACCAGC
GCCGAAGGCG AGACCAGGCC GGCCAAGGAC TGGGTCCCGA GCAACGGACG CACGCTGCGC
CTTCGTAGTG CCAGGGTCAC GAAGTCCAGG CGGGAGATCG GCGCCGTCGT TTGCGTGAAG
GCCAAGGGCA TGAAGGAAGC GTGGCACCTC GCCACGAGCC ACGGCGACAA ACCCGGCTCC
GAGATCGTGG CGCTCTACGC CAGGCGCTTC ACCATCGAGG AGAGCTTCCG CGATCAGAAG
AACCTTCGGT TCGGGATGGG CCTCTCCGAG ACCCGCATCG CCGACCCAGC GCGGCGAGAC
CGCCTGCTCC TCGTCAGCGC GGTCGCAATC GCGCTCCTCA CGATCCTCGG CGCCGCAGGC
GAGGCGCTCG GGCTGGACAA GTGGCTCAAA ACCAACACCG TCAAGCGCCG CACCATCTCG
CTCCTGCGCC AGGGGATGAT GCACTATGCG GCCCTCCCCA AGATGAAGCT CGACATGCTC
GAGCCCCTCA TGGCGAAGTT CGGCGAGATG CTCCGCGCCC AGCGCGTCTT CCGCGAGGTC
TTCGGCCTCA TATGA
 
Protein sequence
MSDQPLDRDR ITKLVSTIFA EDLHAKRVAS LAGAAVGVLE GAALGIHAIG NSLAVAEGLK 
SKHAVKQVDR MLSNEGIPVW RLFGSWVPCV VGDRLEIVVA LDWTDFDEDD QSTIALSMIT
SHGRATPLLW KTVMKSELKG WRNEHEDVLL ERFREVLPEG VKVTVLADRG FGDQALYELL
KDQLGFGFIV RFRGVVKVTS AEGETRPAKD WVPSNGRTLR LRSARVTKSR REIGAVVCVK
AKGMKEAWHL ATSHGDKPGS EIVALYARRF TIEESFRDQK NLRFGMGLSE TRIADPARRD
RLLLVSAVAI ALLTILGAAG EALGLDKWLK TNTVKRRTIS LLRQGMMHYA ALPKMKLDML
EPLMAKFGEM LRAQRVFREV FGLI