Gene Sare_2798 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagSare_2798 
Symbol 
ID5706154 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameSalinispora arenicola CNS-205 
KingdomBacteria 
Replicon accessionNC_009953 
Strand
Start bp3177791 
End bp3178816 
Gene Length1026 bp 
Protein Length341 aa 
Translation table11 
GC content70% 
IMG OID641272254 
Productluciferase family protein 
Protein accessionYP_001537624 
Protein GI159038371 
COG category[C] Energy production and conversion 
COG ID[COG2141] Coenzyme F420-dependent N5,N10-methylene tetrahydromethanopterin reductase and related flavin-dependent oxidoreductases 
TIGRFAM ID[TIGR03559] probable F420-dependent oxidoreductase, Rv3520c family 


Plasmid Coverage information

Num covering plasmid clones20 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones
Fosmid unclonability p-value0.000196172 
Fosmid HitchhikerNo 
Fosmid clonabilitydecreased coverage 
 

Sequence

Gene sequence
ATGAAGCTTG GCTACACCAC CGGCTATTGG TCCGCCGGAC CGCCCGAGGG CGTCACAGCT 
GCCATCGCGG AGGCCGACCG GCTCGGCTTC GACTCGATCT GGGCCGCCGA GGCGTACGGG
TCGGACTGCC TGACCCCACT CGCCTGGTGG GGAGCCAACA CCTCCCGCGT CCGGCTGGGC
ACCAACATCA TGCAGATGGC GGCCCGCACC CCGACCGCCG CGGCGATGGC CGCGCTCACC
CTCGACCACC TCTCGGGCGG CCGGTTCATC CTCGGGCTCG GTGCCTCCGG CCCACAGGTC
GTCGAGGGGT GGTACGGCCA GCCGTACCCG CGACCGCTGG CCCGTACCCG GGAGTACATC
GAGATCGTTC GTACCGTCCT CGCCCGTACC GGACCGGTCG AGCACGACGG AGCGTTCTTC
CAACTCCCGT ACCTCGGCGG CACCGGCCTG GGTAAGCCGC TGAAGTCCAC CGTCCACCCG
CTGCGCGCCG ACATCCCAAT CTTCCTCGCC GCCGAGGGGC CGAAGAACGT GGCCCTGGCC
GCCGAGATCG CCGACGGCTG GCTGCCGTTG TTCTTCTCCC CCAAGGCGGA CAGTTTCTAC
CGTGCCGCAC TCGCCGAGGG CTTCGCCCGG CCTGGTGCCC GCCGTGACAT GGACGCGTTC
GAGGTCGCCG CGACCGTGCC GATCGTCGTC CACGACGACA TCGAGGCAGC CGCCGACCGG
CTCCGGCCGT TCGTCGCGCT GTACGTGGGG GGCATGGGGG CCAAGTCGGC CAACTTCCAC
CGCGACGTCA TCGCCCGCCT CGGGTACGAA CGAGACTGTG ACGTCATCAC CGAGGCATAC
CTGGCAGGTG ACAAAAGGGG GGCTGCCGCC GCCGTACCGA CCGCGCTGGT GGAGGACATC
GCGCTGATCG GCCCGGTCGC CAAGGTCAGG GACGAGTTGC AGGGATGGCG CGAGAGTGTG
GTCACCACCC TGCTCGTCCA GGGCAACTCC CGGCAGCTAC GCCAGATCGC CGAGTTGATG
AGCTGA
 
Protein sequence
MKLGYTTGYW SAGPPEGVTA AIAEADRLGF DSIWAAEAYG SDCLTPLAWW GANTSRVRLG 
TNIMQMAART PTAAAMAALT LDHLSGGRFI LGLGASGPQV VEGWYGQPYP RPLARTREYI
EIVRTVLART GPVEHDGAFF QLPYLGGTGL GKPLKSTVHP LRADIPIFLA AEGPKNVALA
AEIADGWLPL FFSPKADSFY RAALAEGFAR PGARRDMDAF EVAATVPIVV HDDIEAAADR
LRPFVALYVG GMGAKSANFH RDVIARLGYE RDCDVITEAY LAGDKRGAAA AVPTALVEDI
ALIGPVAKVR DELQGWRESV VTTLLVQGNS RQLRQIAELM S