Gene Ppha_0243 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagPpha_0243 
Symbol 
ID6461632 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism namePelodictyon phaeoclathratiforme BU-1 
KingdomBacteria 
Replicon accessionNC_011060 
Strand
Start bp234715 
End bp235746 
Gene Length1032 bp 
Protein Length343 aa 
Translation table11 
GC content51% 
IMG OID642726537 
Productvon Willebrand factor type A 
Protein accessionYP_002017196 
Protein GI194335402 
COG category[R] General function prediction only 
COG ID[COG4245] Uncharacterized protein encoded in toxicity protection region of plasmid R478, contains von Willebrand factor (vWF) domain 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones
Plasmid unclonability p-value0.0131225 
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clonesn/a 
Fosmid unclonability p-valuen/a 
Fosmid Hitchhikern/a 
Fosmid clonabilityn/a 
 

Sequence

Gene sequence
ATGTGTTTTG CCTGGCCAGA TAATATTGGA TATTTACTTT TTTTGATGCC CCTTGCCGTG 
ATTTTGGGGT ATGGAGTCGT AAGGCAGCTT CATGCACGCG AAGCCGTTTT TGGCCCTGCG
CTGATTGATG CTATGATGGG ACGCCTGTCT TTGAGAGTTC TGGTTGTAAA AAAACTGTTG
ATTTTTTGTG GTATCGCGTT GCTGCTCTTT GCGTTGGCGG GTCCCCGGTT TTGCAGTGGA
GGCCGACCGG TGTTGCGCAA AGGTGCCGAT ATCGTTTTTA TGCTTGATGT TTCCCGGAGT
ATGAGGGCAA GAGATGTGCT TCCTGACCGT CTCGGGCAGG CGAAGCAGGA GATAACAAGT
ATCAGTCGTG CTGTCACTGG CGGACGGATG TCTATTCTTC TTTTTGCTGC CAGTCCACTG
GTTCAGTGCC CCCTTACGAC GGATCGGGAT GCTTTCGATG CTCTGCTTGG CATGGCTTCA
CCCGATCTGA TCGAAGAGCA GGGTACCTCT TTCCGTGCGG CGTTTGAGCT TGCCGGACGA
CTTCTTGAAC CGACATTGGA GGATCGAATG GCATCAGGGG TAAAAGGAGA GAAGATTGTG
GTGCTGCTGA GTGATGGGGA GGATCATACC GGTGAGGTTC GGAGTGCAGT CCAGCAGTTG
AAAAAAGCGA ATGTTCATTT GTTTGTGATA GGAGTTGGTA TGCGTCAGCC TGTTGTGATT
CCGTTGGATG ATGCCGGGGA AGGAGTCAAA CGGGATGAGC ATGACAGGGT GATCATGAGC
AGTTTCAGAC CTGAATTCTT ACAGATGCTG GCCCGTGAAG CTGCCGGGTT TTATTTTCGA
AGCAGTGCCG AACATGCCGT TTATAAGGAG GTTTCTGAAA GCATTAACCG TATTGCCTCC
GCTTCCCGAT GGGTGATGGA GCCTGGTGAG CGTGAACCGC TTTATCGTTA TTTTGTTGCG
GCAGGACTTT TTTTACTGCT CACTGAAACT ATGGTTGGAA GAGCTGCAGG AAAGAGGCGG
AGCTGTTCCT GA
 
Protein sequence
MCFAWPDNIG YLLFLMPLAV ILGYGVVRQL HAREAVFGPA LIDAMMGRLS LRVLVVKKLL 
IFCGIALLLF ALAGPRFCSG GRPVLRKGAD IVFMLDVSRS MRARDVLPDR LGQAKQEITS
ISRAVTGGRM SILLFAASPL VQCPLTTDRD AFDALLGMAS PDLIEEQGTS FRAAFELAGR
LLEPTLEDRM ASGVKGEKIV VLLSDGEDHT GEVRSAVQQL KKANVHLFVI GVGMRQPVVI
PLDDAGEGVK RDEHDRVIMS SFRPEFLQML AREAAGFYFR SSAEHAVYKE VSESINRIAS
ASRWVMEPGE REPLYRYFVA AGLFLLLTET MVGRAAGKRR SCS