Gene Avi_2102 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagAvi_2102 
Symbol 
ID7386903 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameAgrobacterium vitis S4 
KingdomBacteria 
Replicon accessionNC_011989 
Strand
Start bp1727106 
End bp1728326 
Gene Length1221 bp 
Protein Length406 aa 
Translation table11 
GC content60% 
IMG OID643651312 
Productaminohydrolase protein 
Protein accessionYP_002549507 
Protein GI222148550 
COG category[F] Nucleotide transport and metabolism
[R] General function prediction only 
COG ID[COG0402] Cytosine deaminase and related metal-dependent hydrolases 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.927814 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGGCGATTG ATTTCGTGTT GCGCCGCGCC AGGTTGCCAT TGTCGGCGCA ACCGCTGGAC 
ATCGCCTTTG AGGCGGGGCG GATCGTTGCG CTGGAAGCGG ATTTCCGTTG TGATGCGCCG
CAGGAGGATG CGGCAGGCCG GTTGGTCTGC GCCGGGTTGA TCGAAACCCA TCTGCATCTC
GACAAGGCAG GGATCATCGG GCGCTGCCGG GTGTGTAGCG GAACGCTGGC GGAAGCCGTG
TCGGAGACCT CGAAAGCCAA GCAGGCCTTT ACCGAGGAAG ATGTCTATGC CCGCGCCGCC
GATGTCGTGG AACGGGCCAT CGTCCAAGGC ACGACCCGGA TCAGGACCTT CGTGGAAGTC
GATCCACGCG CCGGTTTCCG ATCGTTTTCG GCGATCCGCA AGCTGAAGGC CGATTACGCC
CACCTGGTCG ATATCGAAAT CTGCGCCTTT GCCCAGGAAG GGTTGACCAA TGAGCCGGAA
ACCGAGCGGA TGCTGGAAAT CGCCCTGTCG CAAGGGGCCG ATCTGGTTGG CGGCTGCCCT
TACACCGATC CAAGGCCCGC CGAGCATATT TCCCGAATTT TCGAGCTCGC GCAGCGCTTC
GATGTGCCTG TCGATTTTCA CCTTGATTTC GATCTCGATC CTTCCGGGTC CAACCTGCCG
ACGGTCATTG CCCAGACGCT GGCGCGTGGC TATCAAGGCA AGGTCTCTGT CGGCCATGTC
ACCAAGCTTT CCGCAATTTC TCCCGACGAA GTGGAGCGGG TGGCAAAGCA ATTGGCCGAG
GCGGGTATTA CCGTGACGGT TCTGCCTGCC ACCGACCTGT TTCTGACCGG TCGGGATATC
GATCATCTTT GCCCAAGGGG AGTTGCTCCG GCGCATCTTC TGGCCCGCCA GGGGGTGAAT
GTCACCATCT CCACCAATAA TGTTCTCAAC CCATTTACGC CCTTTGGCGA TGTCTCGCTG
ATGCGCATGG CCAATCTCTA CGCCAATGTT GCCCAACTGG CGACGCCTGC GGATCTGAAC
CAGGTCTTCG AGATGATTAC CCGCTATCCG GCCCGGCTGA TGGGGCTGGA TGAACAGCTG
AAGGTTGGGG CTGCGGCTGA TCTTGTCCTT TTTGATGCCG TCTCCGGTGC CGAGGCCGTG
GCGACGATTG CACCGGCGGT GACTGGTTGG AAAAACGGTG TGAAGACGTT CGAACGCAAG
CCGCCGCAGC TCTATCGGTA A
 
Protein sequence
MAIDFVLRRA RLPLSAQPLD IAFEAGRIVA LEADFRCDAP QEDAAGRLVC AGLIETHLHL 
DKAGIIGRCR VCSGTLAEAV SETSKAKQAF TEEDVYARAA DVVERAIVQG TTRIRTFVEV
DPRAGFRSFS AIRKLKADYA HLVDIEICAF AQEGLTNEPE TERMLEIALS QGADLVGGCP
YTDPRPAEHI SRIFELAQRF DVPVDFHLDF DLDPSGSNLP TVIAQTLARG YQGKVSVGHV
TKLSAISPDE VERVAKQLAE AGITVTVLPA TDLFLTGRDI DHLCPRGVAP AHLLARQGVN
VTISTNNVLN PFTPFGDVSL MRMANLYANV AQLATPADLN QVFEMITRYP ARLMGLDEQL
KVGAAADLVL FDAVSGAEAV ATIAPAVTGW KNGVKTFERK PPQLYR