Gene SNSL254_A2449 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagSNSL254_A2449 
Symbol 
ID6486362 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameSalmonella enterica subsp. enterica serovar Newport str. SL254 
KingdomBacteria 
Replicon accessionNC_011080 
Strand
Start bp2364897 
End bp2365955 
Gene Length1059 bp 
Protein Length352 aa 
Translation table11 
GC content61% 
IMG OID642737786 
Productregulatory protein ada 
Protein accessionYP_002041527 
Protein GI194444615 
COG category[F] Nucleotide transport and metabolism
[L] Replication, recombination and repair 
COG ID[COG0350] Methylated DNA-protein cysteine methyltransferase
[COG2169] Adenosine deaminase 
TIGRFAM ID[TIGR00589] O-6-methylguanine DNA methyltransferase 


Plasmid Coverage information

Num covering plasmid clones10 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones53 
Fosmid unclonability p-value0.0458356 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGAAAAAAG CGTTACTTAC CGATGATGAA TGCTGGCTGC GGGTGCAGGC GCGCGATGCC 
AGCGCGGATG GGCGTTTCGT TTTTGCGGTG CGAACCACCG GCGTTTTTTG CCGCCCTTCT
TGTCGCTCGA AGCGGGCGTT ACGTAAAAAT GTTCGCTTTT TTGCCAACGC GCAGCAGGCG
CTGGACGCCG GTTTTCGCCC CTGCAAGCGC TGTCAGCCGG ATAATGCGCG CGCGCAGCAA
CGGCGGTTGG ATAAGATTGC CTGCGCCTGC CGTTTACTTG AGCAGGAGAC GCCGGTAACG
CTGGCGTCTC TGGCGCAGGC GGTGGCGATG AGCCCGTTTC ATCTGCACCG TTTGTTTAAA
GCGAGCACCG GTATGACGCC GAAAGGGTGG CAGCAGGCGT GGCGCGCCCG GCGGCTGCGT
GAGGCGTTGG CGAAAGGAGA GCCGATCACG GCGGCTATTT ATCGCGCCGG CTTCCCGGAT
AGCAGTAGCT ACTACCGTCA TGCCGACCAG ACGCTGGGCA TGACGGCAAA ACAGTTTCGC
AAAGGCGGCG ATAATGTCTC CGTTCGCTAT GCGCTGACGG ACTGGGTTTA CGGACGGTGC
CTGGTGGCGG AGAGCGAGCG GGGGATTTGC GCGATTCTCC CCGGTGATAG CGACGACGCG
CTACTGGCTG AATTACACAC CCTGTTCCCG GCGGCCCGCC ACGAACCTGC TGACGCGCTT
TTTCAGCAAC GGGTGCGGCA GGTTGTCGCG GCTATCAACA CACGCGATGT GCTGCTCTCG
TTGCCGCTGG ATATCCAGGG AACCGCGTTT CAACAGCAGG TCTGGCAGGC GTTATGCGCG
ATTCCCTGCG GCGAAACCGT AAGCTATCAA CAGCTTGCCG CGACTATCGG CAAACCCACG
GCAGTACGCG CGGTCGCCAG CGCGTGCGGC GCGAATAAAC TGGCGATGGT GATCCCGTGT
CATCGGGTCG TGCGTCGCGA TGGCGCGCTC TCCGGTTATC GTTGGGGCGT GCGTCGAAAA
GCGCAGCTAT TAAAGCGAGA AGCACAAAAA GAGGAGTAG
 
Protein sequence
MKKALLTDDE CWLRVQARDA SADGRFVFAV RTTGVFCRPS CRSKRALRKN VRFFANAQQA 
LDAGFRPCKR CQPDNARAQQ RRLDKIACAC RLLEQETPVT LASLAQAVAM SPFHLHRLFK
ASTGMTPKGW QQAWRARRLR EALAKGEPIT AAIYRAGFPD SSSYYRHADQ TLGMTAKQFR
KGGDNVSVRY ALTDWVYGRC LVAESERGIC AILPGDSDDA LLAELHTLFP AARHEPADAL
FQQRVRQVVA AINTRDVLLS LPLDIQGTAF QQQVWQALCA IPCGETVSYQ QLAATIGKPT
AVRAVASACG ANKLAMVIPC HRVVRRDGAL SGYRWGVRRK AQLLKREAQK EE