Gene Slin_4408 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagSlin_4408 
Symbol 
ID8728168 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameSpirosoma linguale DSM 74 
KingdomBacteria 
Replicon accessionNC_013730 
Strand
Start bp5343394 
End bp5344767 
Gene Length1374 bp 
Protein Length457 aa 
Translation table11 
GC content54% 
IMG OID 
ProductAldehyde Dehydrogenase 
Protein accessionYP_003389188 
Protein GI284039258 
COG category 
COG ID 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones25 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones35 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGTTTACGT CCATAAATCC GTACTCACAG AAACGCATTC AAACTTATCG AGCCGATAGC 
GTCCCTGCTA TCGAGCGAAA ACTTAAACAA GCCGACCGGG CCTTTGCCGA CTGGTCGGCT
TTATCGCTTC CGGAGCGTAC AGATTATTTA CGCAAAGTCG GGCAGTATTT AACCGCCAAC
AAACAACGAT ACGGCGAGTT GATAACGGCC GAAATGGGGA AGACGATGAA AGAAGCCATT
GGTGAAGTCG AGAAATGTGC CGCCACCTGT ACGTTCTACG CCGACCATGC CGATGCTTTT
CTCGCCAACC AGTCCATACA ATCGGACCGT ACCGACGGAC CGGCCGAAAA GAGCTTTGTC
ACCTACCAGC CATTGGGGGT TGTGCTGGCT ATTATGCCGT GGAATTTTCC ATTTTGGCAG
GTCATGCGCT TTGCCATTCC AGGCCTGATT GCTGGCAATG TGGGTTTGCT CAAGCACGCG
CCTAATGTAT TCGGCTGTTC GCTGGCTATT GAAGAAGCTT TTCGGGAATG CGGGTTGCCG
ACGGGCGTTT TTCAATCGCT GTTGGTGGAT GTTCCAGTGG TGGAAGACTT ATTGAAAGAC
AAACGGGTAA AAGCCGTGAC GCTTACGGGC AGCGGGCGCG CCGGTTCGTC GGTTGCGGGT
ATTGCCGGAC GGGAAATCAA AAAGTCGGTG CTGGAACTAG GCGGCTCCGA CGCGCTGATC
GTTCTGGCCG ATGCTGATCT GGACAAAGCG GCCGAGGTAG CCGTGAAGTC GCGCATGCAA
AACGCCGGAC AAAGCTGCAT TGCCGCTAAA CGGTTTATTA TTGAAAAGTC CGTGAAAAAA
GCTTTCACCG AACGGGTGCT GCATCAGATC AGTCAAATCC GGCAGGGCGA CCCTATGGAT
GAAGCGACCA CGATGGGGCC AATGGCGCGG CTCGATCTGG CCGATTCCAT TGAGCGGCAA
TATCAGCAGA CGCTCGCTAA AGGAGCTAAA CGACTGACGG GCGGTGAGCG CGAAGGGTGT
AATGTGCAGC CGATACTGCT GGACAACGTG AAGCCCGGCA TGGCGGCTTT CGACGAGGAA
ACCTTCGGGC CACTGGCCGT TCTGATCGAC GCTAAAGATG AAGTGGACGC TATTCGACTG
GCAAATCACT CTGATTTTGG CCTCGGCTCG GCCATCTGGA CGCAGGATCT GGACAGGGCA
GAGCGGCTGG CCTACCAGAT TCAGGCGGGG TCCGTTTTTA TCAATGGGCT CATGCGCTCC
GATGCCCGCG TGCCTTTTGG TGGTATCAAA ATGTCAGGCT TTGGCCGCGA ATTATCGGAA
GCGGGTATCA AAGAATTTAC GAATGTGAAA ACGATCTGGG TAGAGAAAAG CTAG
 
Protein sequence
MFTSINPYSQ KRIQTYRADS VPAIERKLKQ ADRAFADWSA LSLPERTDYL RKVGQYLTAN 
KQRYGELITA EMGKTMKEAI GEVEKCAATC TFYADHADAF LANQSIQSDR TDGPAEKSFV
TYQPLGVVLA IMPWNFPFWQ VMRFAIPGLI AGNVGLLKHA PNVFGCSLAI EEAFRECGLP
TGVFQSLLVD VPVVEDLLKD KRVKAVTLTG SGRAGSSVAG IAGREIKKSV LELGGSDALI
VLADADLDKA AEVAVKSRMQ NAGQSCIAAK RFIIEKSVKK AFTERVLHQI SQIRQGDPMD
EATTMGPMAR LDLADSIERQ YQQTLAKGAK RLTGGEREGC NVQPILLDNV KPGMAAFDEE
TFGPLAVLID AKDEVDAIRL ANHSDFGLGS AIWTQDLDRA ERLAYQIQAG SVFINGLMRS
DARVPFGGIK MSGFGRELSE AGIKEFTNVK TIWVEKS