Gene Nwi_3053 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagNwi_3053 
Symbol 
ID3676362 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameNitrobacter winogradskyi Nb-255 
KingdomBacteria 
Replicon accessionNC_007406 
Strand
Start bp3315607 
End bp3316851 
Gene Length1245 bp 
Protein Length414 aa 
Translation table11 
GC content61% 
IMG OID637714619 
Productmajor facilitator transporter 
Protein accessionYP_319654 
Protein GI75677233 
COG category[G] Carbohydrate transport and metabolism 
COG ID[COG2814] Arabinose efflux permease 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones27 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones
Fosmid unclonability p-value0.0592542 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGCAGAAAG CGCAGCAGCA TTTCCACGGT CCATGGATCG TCGCCGTCGC CTTCGTGACC 
CTCATGGGGG CATTCGGCCT CAACCTGAGC GGCGGCCAGT TTTTCGCGCC GTTGGCGCAG
GAGTTCGGGT GGGGACTCGC CGCGTTGAGT TCAGCGGCTG CGCTCAACAT GATTGTATGG
GGGATCTTCC AGCCGCTCAC GGGACGGATG ATCGACCGTT TCGGTCCCAA GCCGGTCATC
GTCGGAAGCG CCGCCTTGAT GGGAGCGGCC TTTTTGCTGT CCTCGACGAT TTCAACGGTT
TGGGAGTTCT ATCTCTACTA CGGAGTGCTT GCCGCCATTG GCTTCGCCGG CTGCGGATCG
ATGGCCAATT CCGTCCTCGT TGCGCGATGG TATGTCCGGG GACGAGCAAG GATGCTGGCT
CGAAGCGCAA TGGGCATGAA CATCGGCCAG CTTGTTTTCC TTCCGCTGAC GGGATGGCTT
ATCATCGTCT CTGGTTTTCG CGGGGCGTTT CTTGGGCTCG GCGTTTTGAT GATCGTGGTG
ATTGTCCCGC TCGTTCTGTT CGTTGCGCAC AGTTCGCCGG ACAAGATCAA CCAGGCGCCG
GACGGCGACG AGCTTTCCAC CTTTACGGCG CCGAAATCCG CCTCGCTTTC ACAAGCCGTC
AGAAGCCCGG ATTTTTGGGC AGCGACCTGC GGCTTCGTGA CCTGCGGATA CTCGCTTTAT
CTCGTCGTGA TCCATCTACC GCGTTTCGCC GTCGATCTCG GGGCGAACCT TGCCACGGGC
GGACAGGTTC TCGGCCTTGC CGCTGGCGCA AGCGCGATTT CGATGTGGAC GTGCGGCCAA
CTGGCCGGAC GGGCCGGCAA GAAGAACCTG CTGATCGGCC TCTATCTGGT CCGCGCGCTC
TCGCTTGCCT TCCTCGCCGT CTCAACGGAG GTTTGGCAAC TCTATGCGTT CGCTCTTGTC
TACGGCCTGT CGTCCATGCC CATCATTCCG TTGAAGACCG GACTGATCGG GGACCTCTTC
GGGGCAAACG CCATGGGCAG CATACTCGGG ACCGTATGGT TCCTGCATCA GATCCTCGCA
GCCATCGCCG TATATCTAGG CGGCTATTTG CGGGTCGAGA CCGGAAGTTA CGCGGCGGCC
TTCTGGTCAG CCGCGATTCT GCTCCTGATC GGCGCGGCAT CGACTAGTCT TGTTCGCAGC
CCTGGCGGGC CGGTTCCTCT GAAGCCGAAT CCAGCGCGCA GCTAG
 
Protein sequence
MQKAQQHFHG PWIVAVAFVT LMGAFGLNLS GGQFFAPLAQ EFGWGLAALS SAAALNMIVW 
GIFQPLTGRM IDRFGPKPVI VGSAALMGAA FLLSSTISTV WEFYLYYGVL AAIGFAGCGS
MANSVLVARW YVRGRARMLA RSAMGMNIGQ LVFLPLTGWL IIVSGFRGAF LGLGVLMIVV
IVPLVLFVAH SSPDKINQAP DGDELSTFTA PKSASLSQAV RSPDFWAATC GFVTCGYSLY
LVVIHLPRFA VDLGANLATG GQVLGLAAGA SAISMWTCGQ LAGRAGKKNL LIGLYLVRAL
SLAFLAVSTE VWQLYAFALV YGLSSMPIIP LKTGLIGDLF GANAMGSILG TVWFLHQILA
AIAVYLGGYL RVETGSYAAA FWSAAILLLI GAASTSLVRS PGGPVPLKPN PARS