Gene Caul_2962 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagCaul_2962 
Symbol 
ID5900417 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameCaulobacter sp. K31 
KingdomBacteria 
Replicon accessionNC_010338 
Strand
Start bp3217237 
End bp3218361 
Gene Length1125 bp 
Protein Length374 aa 
Translation table11 
GC content70% 
IMG OID641563459 
Productagmatine deiminase 
Protein accessionYP_001684587 
Protein GI167646924 
COG category[E] Amino acid transport and metabolism 
COG ID[COG2957] Peptidylarginine deiminase and related enzymes 
TIGRFAM ID[TIGR03380] agmatine deiminase 


Plasmid Coverage information

Num covering plasmid clones41 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones13 
Fosmid unclonability p-value0.378671 
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGAGCCGCC TCCTGACCAC CACGCCCCGC GCCGACGGCT TTCACATGCC CGCGGAGTGG 
GAGCCCCACG CCGGTTGCTG GATGTTGTGG CCCGAGCGGT CGGACAACTG GCGAGGCGGC
GCCAAGCCGG CCCAGCACGC CTTCGTCGCC GTGGCCGCCG CCATTGTCCA GGGCGAGCCC
GTCACCGTCT GCGTTTCGCC GGCCCAGTAC GTCATCGCCC GCGAGATGCT GGATCCCGCC
GTGCGCGTGG TGGAGATGAC CAGCAACGAC AGCTGGATCC GCGACTGCGG GCCGACCTTC
GTTATCGACG AGGCCGGACG CGTTCGGGGC GTGGACTGGA AGTTCAACGC CTGGGGCGGC
CTGATCGGCG GCCTCTACTT CCCCTGGGAC CAGGACGACC TGGTTGGCGA GAAGGTTATC
GAGCTGGAGG GCGACGATCG CTATGGCCCC GACTTCATCC TCGAGGGCGG CTCGATCGAC
GTCGATGGCC AGGGCACGGT GTTGGCGACC AAGGAGTGCC TGCTCAATCC CAACCGCAAT
CCCGGCCTCG GCCAGGGCGA GATCGAACAG CGCCTGCGCG ACTATCTGGG CGTCGAGACC
GTGATCTGGC TCGACCAGGG CGTCTATCTC GATGAGACCG ACGGCCACGT CGACAATTTC
TGCCGGTTCG TCGCTCCCGG CGAGGTGGTG CTGACCTGGA CCGACGATCA AGCCGATCCG
CAGTACGAGC GTTCGGCCGC CGCCCTCGCG CGCCTGGGCG CGGCGCGCGA TGCTCGCGGG
CGGTCGCTGA ACATCCACAA GCTGCATCAG CCCGCCCCCG TGATCATCAC CGCCGAGGAG
GCCGCCGGCG TCGACAAGGT CCCCGGGACC CTGCCGCGCG AGGCGGGCGA TCGGATGGCC
GCCTCCTACG TCAACTTCTA TGTCGGCAAC GGCGTCGTGG TGGCGCCCGC CTTCGACGAC
CCCATGGACG CGCCGGCCCA GGCGTTGCTG GCCAAGCTGT TTCCGGGGCG CAGGATCCTG
CCTGTCCCCG CCCGCGAGAT CCTCCTGGGC GGCGGCAATA TCCACTGCAT CACCCAGCAG
GAGCCTCTGG CGCGGGGCGC CCCGATGATC GCCCGGCGCG CCTGA
 
Protein sequence
MSRLLTTTPR ADGFHMPAEW EPHAGCWMLW PERSDNWRGG AKPAQHAFVA VAAAIVQGEP 
VTVCVSPAQY VIAREMLDPA VRVVEMTSND SWIRDCGPTF VIDEAGRVRG VDWKFNAWGG
LIGGLYFPWD QDDLVGEKVI ELEGDDRYGP DFILEGGSID VDGQGTVLAT KECLLNPNRN
PGLGQGEIEQ RLRDYLGVET VIWLDQGVYL DETDGHVDNF CRFVAPGEVV LTWTDDQADP
QYERSAAALA RLGAARDARG RSLNIHKLHQ PAPVIITAEE AAGVDKVPGT LPREAGDRMA
ASYVNFYVGN GVVVAPAFDD PMDAPAQALL AKLFPGRRIL PVPAREILLG GGNIHCITQQ
EPLARGAPMI ARRA