Gene Ndas_4707 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagNdas_4707 
Symbol 
ID9248589 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameNocardiopsis dassonvillei subsp. dassonvillei DSM 43111 
KingdomBacteria 
Replicon accessionNC_014210 
Strand
Start bp5585980 
End bp5587266 
Gene Length1287 bp 
Protein Length428 aa 
Translation table11 
GC content75% 
IMG OID 
Productprotein of unknown function DUF1205 
Protein accessionYP_003682599 
Protein GI297563625 
COG category 
COG ID 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones14 
Plasmid unclonability p-value0.806542 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones18 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGCGCGTGC TCGTCGTCAG CCAGGCGGAG AAGACCCATC TGCTGGGCCT CATCCCGCAG 
GCGTGGGCTC TGCGCGCCGC CGGGCACGAG GTGCGGGTGG CCAGCCAGCC CGCGCTGGTC
CCCGTCGCGG CGCGCACCGG GCTGCCCGCC GTCCAGGTGG GCCGGGACCA CCTCTTCCAC
CAGCTGCTGA CCACGTTGAA GGGCCTGGGC TTCGGCGACA GCCGGGGCTT CGACATGACC
AGGAGCGATC CGGAGGCGCT GGGCTGGGAC TACCTGCTCG ACGGCTACCG CGAGTTCGTC
CGGCTGTGGT GGCATCCGGT CAACACGCCC ATGCTGGACG ACCTCACCGA CCTGTGCCGC
TCGTGGCGCC CCGACCTGGT CCTGTGGGAG CCGACCACCT TCGCCGCGCC GGTGGCGGCC
CGGGCCTCGG GCGCGGCGCA CGTCCGGGTC CTGTGGGGGC TGGACGTGTT CTCCCGCACC
CGCCGCCGGT TCCTGGAGCG CGCCGCCGCG CTGTCCGCCG CCGACCGCGA GGACCCCCTC
GCCGACTGGC TGGAGCGCAG CGCCCGGCGC GTGGGCGCCG ACTTCTCCGA GGACCTGGTC
CGGGGCCAGG CCACCCTGGA CCCCTATCCG CCGGGTGTGC GCCTGGACCC CGAGGAGGGC
GTGCGCCACA TCCCCCTGCG CTACGTGCCC TACAACGGGA CCGCGGTGGT GCCCGACTGG
CTGCGCTCCC CGGGCGGGCG CAGGCGGGTC TGCCTGACCC TCGGCTCGGC CGTGCCGGAG
AAGTTCGACG ACCGCTACCG GCTGCCCCTC GCGGAGCTGC TGGAGTCGGT CGCCGGACTG
GACGTGGAGG TCGTCGCCAC CCTGTCCGCC GAGCAGAGCG CGCGGGCCGG AACCCTGCCG
GACAACGTCC GGGTGGTGGA GCACGTGCCG TTGCACGCGC TCATGCCCCA CTGCGACGCG
GTGGTCCACC ACGGGGGAGC CGGGACCTTC TGCACCGCGG TGTTCCACGG TGTCCCGCAG
CTCGTCCTCC CCGAGTTCTC CATGGCCCAG TACGTCTTCG ACGAACCGCT GCTCGCCGAG
CGGATCACCG GGTTGGGGGC GGGCCTCGCG CTGGCGGGGG CCGGGATGAC CGGCGGGGAG
GTCGCCCTCC AGGTCGGGCG CCTGCTCGAC GAGCCCCGCT TCGCCGAGGG CGCGCGCGTG
CTCCGCGACA GGGCGCACGG GATGACCAGC CCGGCCGGAC TCGTCCCGGT GCTGGAGGAG
CTGGCCGCCG AGGGCCGAAG CGCCTGA
 
Protein sequence
MRVLVVSQAE KTHLLGLIPQ AWALRAAGHE VRVASQPALV PVAARTGLPA VQVGRDHLFH 
QLLTTLKGLG FGDSRGFDMT RSDPEALGWD YLLDGYREFV RLWWHPVNTP MLDDLTDLCR
SWRPDLVLWE PTTFAAPVAA RASGAAHVRV LWGLDVFSRT RRRFLERAAA LSAADREDPL
ADWLERSARR VGADFSEDLV RGQATLDPYP PGVRLDPEEG VRHIPLRYVP YNGTAVVPDW
LRSPGGRRRV CLTLGSAVPE KFDDRYRLPL AELLESVAGL DVEVVATLSA EQSARAGTLP
DNVRVVEHVP LHALMPHCDA VVHHGGAGTF CTAVFHGVPQ LVLPEFSMAQ YVFDEPLLAE
RITGLGAGLA LAGAGMTGGE VALQVGRLLD EPRFAEGARV LRDRAHGMTS PAGLVPVLEE
LAAEGRSA