Gene Avi_2101 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagAvi_2101 
Symbol 
ID7386902 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameAgrobacterium vitis S4 
KingdomBacteria 
Replicon accessionNC_011989 
Strand
Start bp1725811 
End bp1727010 
Gene Length1200 bp 
Protein Length399 aa 
Translation table11 
GC content61% 
IMG OID643651311 
Producthydrolase 
Protein accessionYP_002549506 
Protein GI222148549 
COG category[F] Nucleotide transport and metabolism
[R] General function prediction only 
COG ID[COG0402] Cytosine deaminase and related metal-dependent hydrolases 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGAGCGCCA CTCTCGTCGT TTTCAATGCC CGCAATGGTC TCGGTGATCC CGTTGATATC 
GTTATTGCCG GTCATGCCAT TGCCGCTATC GGCCCAGCGG CGGGTGAGGG CGTTTCAACG
GAAAAACCGC GCATCGATGC TCGAGGCGGT CTGGTCCTGC CGGGCCTGGT CGATGGTCAT
GTACATCTGG ATAAAACGCT GATCGGCATG CCGTTCATTC CCCATATTCC CGGAGGCACG
GTCGCCGAGC GGATCCGGGC CGAAAAAGCG CTCCGCCGCT CGCTGCCTTT GCCGGTCGAG
GTGCGTGGGG CTAAGCTGCT GGAGAAGATG GCCACCTATG GCACCGTCGC CTGCCGTAGC
CATGCCGATA TCGATACGGA AGTCGGGTTG GCAGGGCTAG AAGCCATATT GTCCCTGAAG
CAAAGCCATG CTCATCTGGT CGATATTCAG ACAGTGGCGT TTCCACAGTC CGGCGTGCTG
GCCGATCCCG GCACTGCTGA TCTGCTGGAG CAGGCGGTCA AGGCGGGTGC GGACCTGATC
GGTGGTCTTG ATCCAGCCGG GATCGACGAT GACATTACCG GCCACCTGAA CGCGATTTTT
GCTATCGCCG GACGTCACGG CGTGGGGGTG GATCTGCATC TACACGATCC GGGCCCGCTC
GGCGCGTTCG AAATCCGCCA GATCGCCAAA CGGGCTTTGG CGCAGGGACT GCAAGGCAAA
TGCGCCGTCA GCCATGCCTA TTGCCTGGGC GCTCTGGATG ATATGGATTT CGGGCGCACA
GCGGAAGCGC TGGCGCGCGC CGATGTGGCG ATCATGACCA CCGGCCCCGG CGATACCAGC
ATGCCGCCGA TCAAGCGGCT GAAAGCTGCG GGCGTGCGGG TGTTTTCCGG CAATGACAAT
ATCCGCGATG CCTGGTCGCC GCTCGGCAAT GGCGATCTTT TGGAGCGGGC GAGCATTCTC
TGCGACCGGC AGAACTTTCG CGCCGATGCC GACCTTGAAC ATGCTTTCGC GCTTGTCAGT
ACCCTCTCGG CTGAGGTTTT AGGACGCAGC AATGCGACAC TCGGCAAAGG GTGTCCCGCC
GATTTCATCA TTCTGCCGGT TGCCTCGATT GCCGAAGCAG TGGCGGCGCG TCCGATGGAG
CGTATGGTGT TCAAGGCAGG TGTGCTGGTT GCCAGCAATG GCAATCTCGT CGCCTCATGA
 
Protein sequence
MSATLVVFNA RNGLGDPVDI VIAGHAIAAI GPAAGEGVST EKPRIDARGG LVLPGLVDGH 
VHLDKTLIGM PFIPHIPGGT VAERIRAEKA LRRSLPLPVE VRGAKLLEKM ATYGTVACRS
HADIDTEVGL AGLEAILSLK QSHAHLVDIQ TVAFPQSGVL ADPGTADLLE QAVKAGADLI
GGLDPAGIDD DITGHLNAIF AIAGRHGVGV DLHLHDPGPL GAFEIRQIAK RALAQGLQGK
CAVSHAYCLG ALDDMDFGRT AEALARADVA IMTTGPGDTS MPPIKRLKAA GVRVFSGNDN
IRDAWSPLGN GDLLERASIL CDRQNFRADA DLEHAFALVS TLSAEVLGRS NATLGKGCPA
DFIILPVASI AEAVAARPME RMVFKAGVLV ASNGNLVAS