Gene WD1033 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagWD1033 
Symbol 
ID2738829 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameWolbachia endosymbiont of Drosophila melanogaster 
KingdomBacteria 
Replicon accessionNC_002978 
Strand
Start bp993355 
End bp994584 
Gene Length1230 bp 
Protein Length409 aa 
Translation table11 
GC content35% 
IMG OID637173188 
Productpermease, putative 
Protein accessionNP_966757 
Protein GI42520842 
COG category[G] Carbohydrate transport and metabolism 
COG ID[COG2271] Sugar phosphate permease 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.187473 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGTTGCAGA GGAATTTTTT AATCTGGTTG TTGACGTCAC TGTTTTATGC ATACCAATAT 
ATATTACGCG TAATTCCAAA CATAATTGCG CCTGAATTAA TAACAAAATT TAACATAAGT
ATTGCAGATG TTGGCCAGTT TGGTGGGTTA TACTATGTAG GCTATACGCT AGCTCATATA
CCTATTGGTC TTTTCCTTGA TAGATTTGGA CCAAAGTTTG TTTTACCTGC TTGTACCGTT
TTGACATTTA CCGGAACACT ACCGCTTATA TGCTTTGATG AGTGGAGTTA TTCAATAATT
GGAAGAATAA TTGTGGGAAT TGGGTCATCC GCTTCGGCAA TTGGAATCTT TAAAGTTGCA
AGCATGTATT TTGCGCAAGA AAAATCAGCA AGAATGGCCA GCTTATCTAT AATTATAGGA
ATTTTGGGAG GGATATGTGG CGGGTTACCT TTAGACTTCT TACTCGATAA ATTTGGCTGG
AATTATGTTA TCTATACATT CTCAGCATTT GGGTGTTTAC TTGCTCTGTT GCTGTTCTTA
GTAACGCCTG AAAGCAATGC TCAACAGGAA AAAGTTAGCA TTAGAGACTT AAAAAACATA
CTTTTCAACA AGCATATTAT TCTAATTAGC TTTTTTGGTG GACTCATGGT CGGTCCAATG
CAAGGTTTTG CCGATGGTTG GGTGAAAGCA TTCTTTTTTG AAGTATATAA AATGAATGAA
GACTTGGCAT CTTCTCTCTC TTCCGTAATA TTGATAGGAA TGTTAACAGG ATCATTCTCT
TTGGCTTATT TATTGGAAAA ATATAAAAAT AAGCATTATG AAGTAATAAT TGCATGCTCA
TTTGCAATGA TCGCTAGTTT TCTTTTGCTT TTTACTCAGA TTGGTGGCTT GTATGTTGTA
TTACCTACGC TTTTTATTAT TGGCTTCACA TCTGGATATC AGGTGGTTAC CATTTACAAA
GCGATAAGTT ATGTAAATAA TAACTTAGTA GGCCTGGCTA CAGCTGTGTC AAACATGATA
GTTATGGTTT TTGGCTATTT TTTTCACACT GGGATCGCAA AAATAGTAGA TTTGTGTTGG
GATAGAACAA TAATACAAGG AAATCCTGTG TATGGTGCTG AATTGCTGAT AAAAGCAACA
TCAGTTATTC CTGTATGTTT GCTGGTGGCT GTTTTTGGGC TCTTATGGTT AAAAAATAAA
GATTTTAGAG AAGTTGATTG TAATAAGTAG
 
Protein sequence
MLQRNFLIWL LTSLFYAYQY ILRVIPNIIA PELITKFNIS IADVGQFGGL YYVGYTLAHI 
PIGLFLDRFG PKFVLPACTV LTFTGTLPLI CFDEWSYSII GRIIVGIGSS ASAIGIFKVA
SMYFAQEKSA RMASLSIIIG ILGGICGGLP LDFLLDKFGW NYVIYTFSAF GCLLALLLFL
VTPESNAQQE KVSIRDLKNI LFNKHIILIS FFGGLMVGPM QGFADGWVKA FFFEVYKMNE
DLASSLSSVI LIGMLTGSFS LAYLLEKYKN KHYEVIIACS FAMIASFLLL FTQIGGLYVV
LPTLFIIGFT SGYQVVTIYK AISYVNNNLV GLATAVSNMI VMVFGYFFHT GIAKIVDLCW
DRTIIQGNPV YGAELLIKAT SVIPVCLLVA VFGLLWLKNK DFREVDCNK