Gene Aazo_3449 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagAazo_3449 
Symbol 
ID9341253 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism name'Nostoc azollae' 0708 
KingdomBacteria 
Replicon accessionNC_014248 
Strand
Start bp3519896 
End bp3521065 
Gene Length1170 bp 
Protein Length389 aa 
Translation table11 
GC content37% 
IMG OID 
Productcysteine desulfurase 
Protein accessionYP_003722206 
Protein GI298492029 
COG category 
COG ID 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones13 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGTCCAGTC GTCCTATCTA CCTCGATTGT CACGCTACCA CACCCATAGA TGAACGGGTA 
CTAAATGCAA TGATTCCTTA CTTTACAGAA AAGTTTGGTA ATCCAGCTAG TATTAGTCAT
GTTTATGGTT GGGAATCAGA AGCCGCTGTT AAACAATCCA GAGATATTTT AGCAACTGCT
ATTAATGCTA ACCCGGAAGA AATTGTCTTT ACTAGTGGTG CAACAGAAGC TAATAATTTA
GCCATCAAAG GTGTAGCAGA AGCTTACTTT GCCAAAGGTC AACATATTGT TACAGTTGCA
ACCGAACATA AAGCGGTTTT AGAGCCTTGT GAATATTTAG AAAGCATGGG TTTTGAAATT
ATGGTTCTTC CAGTTAATCA AGATGGTCTA ATTGATTTAG AGCAATTAGA AAAAACCTTG
CGTCATGATA CAATTTTAGT ATCAGTCATG GCTGCAAATA ATGAAATCGG AGTTTTACAA
CCCTTAGATA AAATCGGTAA AATGTGCCGT CAAAAAGAAA TTATATTTCA TACAGATGCA
GCTCAAGCCA TTGGTAAAAT TCCCTTAGAT GTAGAAGCAT TAAATATTGA TTTAATGTCC
TTAACAGCCC ATAAAGTCTA TGGACCAAAA GGTATTGGTG CTTTATATGT TCGCAGACGC
AACCCCAGAA TTAAACTAGC AGCACAACAG CATGGGGGTG GCCATGAAAG AGGAATGCGT
TCTGGGACAT TATATACACC CCAAATCGTT GGTTTTGCTA AAGCTGTAGA AATTGCTTTA
GCAGAACAAG AAACTGAAAA TCAACGCTTA ACAGAACTGC GGGAAAGATT GTGGAAACAG
TTATCTACTC TGGAGGGAAT TTATATTAAT GGACATCCCC AAAAACGGTT GGCAGGAAAT
TTAAATATTA GTCTTGAAGG TGTAGATGGT GCTGCACTTT CTTTAGCTTT ACAACCAATG
GTAGCAGTAT CTTCTGGTTC TGCTTGTTCC TCAAATAATG TTGCACCTTC CTATGTGCTG
ATAGCTTTAG GTCATCCAGA AAAATTAGCT TATGCTTCCG TGCGATTTGG AATGGGTAGA
TTTAATACTG TTGAAGAAAT AGATAAAGTA GCAGAACATT TCATTACTAC TGTGAAAAGT
TTAAGAAGTA CTTCAGTAGT CATTTGTTAG
 
Protein sequence
MSSRPIYLDC HATTPIDERV LNAMIPYFTE KFGNPASISH VYGWESEAAV KQSRDILATA 
INANPEEIVF TSGATEANNL AIKGVAEAYF AKGQHIVTVA TEHKAVLEPC EYLESMGFEI
MVLPVNQDGL IDLEQLEKTL RHDTILVSVM AANNEIGVLQ PLDKIGKMCR QKEIIFHTDA
AQAIGKIPLD VEALNIDLMS LTAHKVYGPK GIGALYVRRR NPRIKLAAQQ HGGGHERGMR
SGTLYTPQIV GFAKAVEIAL AEQETENQRL TELRERLWKQ LSTLEGIYIN GHPQKRLAGN
LNISLEGVDG AALSLALQPM VAVSSGSACS SNNVAPSYVL IALGHPEKLA YASVRFGMGR
FNTVEEIDKV AEHFITTVKS LRSTSVVIC