Gene Acid345_0962 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagAcid345_0962 
Symbol 
ID4072950 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameCandidatus Koribacter versatilis Ellin345 
KingdomBacteria 
Replicon accessionNC_008009 
Strand
Start bp1220176 
End bp1221285 
Gene Length1110 bp 
Protein Length369 aa 
Translation table11 
GC content61% 
IMG OID637982969 
Productglycosy hydrolase family protein 
Protein accessionYP_590039 
Protein GI94967991 
COG category[R] General function prediction only 
COG ID[COG4225] Predicted unsaturated glucuronyl hydrolase involved in regulation of bacterial surface properties, and related proteins 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones19 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones11 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGAAGCGAG CCATCATCCT AGCCGCTGTC TTGCTCTGCA TGCCGGTCTT TGCCGCGCAA 
GCCCAGCCCG CACCGTCCGC CACCGAAGTG CTGGCGACGA TGGAGCGCGT CGCCGACTGG
CAACTCGCGC ATCCCTCGTC TCACGCCACA ACCGAGTGGA CGCAAGCCGC CGGATACGCC
GGCATGATGG CGCTCGCGAA TCTTTCGAAG GGCCCCAGGT ACCGTGAAGC CGTGCGCTCG
ATGGGGGAAG CCAACGCCTG GAAGGCCGGC CCGCGCACGT ATCACGCCGA TGACTTAGCG
GTTGGCCAGA CCTACGAAGC ACTCTACGTC TTCTACAAAG ATCCCAAGAT CATCGCGCCC
CTTCGCGAGC GCCTCGACTT CATCCTCGCG ACCCCGTCCA AAGTCCAGTC GCTCGACTTT
CAGCAGAAAT ATGATCAGGT GAGCGAACTA TGGTCGTGGT GCGATTCGCT TTTCATGGCG
CCGCCGGTGT GGGTCGAAAT GTCGGCGATC ACTGGCGATC CGCGCTACCG CGAGTTCGCG
ATCAAAGACT GGTGGCGCAC CACGGATTTC CTTTACGATC CCGCCGAGCA TCTTTATTAC
CGCGACAGCA CGTATTTCAA TCGCCGCGAA GCCAACGGGC AGAAGGTGTT CTGGAGTCGC
GGCAACGGCT GGGTGATGGC CGGACTTGTT CGCGTGCTTG ATGCTTTGCC CCCGCAGGAT
CCATCGCGGC CGCGCTTCGA GAAGCTCTAT CGCGACATGG CGCAAGCCGC GCTTAAAGCA
CAGCAGCCCG ACGGTCTTTG GCACGCCAGT CTTCTCGATC CGCAGAGCTA TCCACTCAAA
GAGACCAGCG GCTCTGGTTT CTTTACCTAT GCTCTGGCGT GGGGAGTGAA CCACAAGCTG
CTCGACCGCG CTGAATACGA GCCCGCTGTC GTAAGAGCAT GGAGCGCGCT GGTGGCTTGC
GTGGATAAAG ACGGCAAGCT CACGCACGTC CAGCCCATTG GCTCCGATCC CAAATCGTTC
GCCGAGGATT CGACTGAGGT ATATGGCGTC GGGGCATTCC TGCTCGCAGG CAGCGAGGTC
TATCGCATAG CTGAGAAGGG AAAGCACTGA
 
Protein sequence
MKRAIILAAV LLCMPVFAAQ AQPAPSATEV LATMERVADW QLAHPSSHAT TEWTQAAGYA 
GMMALANLSK GPRYREAVRS MGEANAWKAG PRTYHADDLA VGQTYEALYV FYKDPKIIAP
LRERLDFILA TPSKVQSLDF QQKYDQVSEL WSWCDSLFMA PPVWVEMSAI TGDPRYREFA
IKDWWRTTDF LYDPAEHLYY RDSTYFNRRE ANGQKVFWSR GNGWVMAGLV RVLDALPPQD
PSRPRFEKLY RDMAQAALKA QQPDGLWHAS LLDPQSYPLK ETSGSGFFTY ALAWGVNHKL
LDRAEYEPAV VRAWSALVAC VDKDGKLTHV QPIGSDPKSF AEDSTEVYGV GAFLLAGSEV
YRIAEKGKH