Gene Rcas_3086 details

Gene Information       Plasmid Coverage information       Fosmid Coverage information       Sequence       

Gene Information

Locus tagRcas_3086 
Symbol 
ID5540582 
TypeCDS 
Is gene splicedNo 
Is pseudo geneNo 
Organism nameRoseiflexus castenholzii DSM 13941 
KingdomBacteria 
Replicon accessionNC_009767 
Strand
Start bp3998794 
End bp4000791 
Gene Length1998 bp 
Protein Length665 aa 
Translation table11 
GC content61% 
IMG OID640895205 
Productpolymorphic outer membrane protein 
Protein accessionYP_001433158 
Protein GI156743029 
COG category 
COG ID 
TIGRFAM ID 


Plasmid Coverage information

Num covering plasmid clones11 
Plasmid unclonability p-value
Plasmid hitchhikingNo 
Plasmid clonabilitynormal 
 

Fosmid Coverage information

Num covering fosmid clones42 
Fosmid unclonability p-value
Fosmid HitchhikerNo 
Fosmid clonabilitynormal 
 

Sequence

Gene sequence
ATGAACACCT CACGACCACC CAGTCGCTGG CGTCGTCTGC TGGCGCATAT CTGCCTGACC 
ATCCTGTGTG TCAGCGCGAC AGTTCCGCCA CTGCCCGGTA TGGCATCGGC GCCTGCCGGC
GCTCCTGGCG CGCCTCCCGC TGCCTGCACC CCTCCGATTG TGCCGGTGAC GCTGGTCAAT
CCGACGGTCA TCACCAGTTG CACGCAGGCG AACCTTCAGG CGGCGCTTGC GACCGGCGGG
CATATCACCT TCGACTGCGG ACCCCACCCG GTCACGATTC CGATCACTTC GCCACTGGTC
ACCTCAGCGA CCCGCGATAT TGTGCTGGAT GGTAAAGGAC TGATCACGCT CGATGGCGGC
GGTGTCACCC GTATTCTCGA AAAACCGTTT ACCCCGGGTT CGCACATTGA TAAGACCTCC
GGCAACGATC TGGTCATCCA GAATATGCGC TTTATCAATG GGCGCGCGCC GGCCGCCACG
AAAACGCAGG ACGACAAAGC GCGCGGTGGG GCGCTGTGGG TGACGAGTCC GGGAACACGC
CTCCATATCA TCAACTCGAT CTTCGAGAAC AACCGCACGA CCAGCATGAC CGACGAGGAC
AATCAAGGCG GTGCGATCTA TGCCGGCAAT ATCTACGAAA CGGTCATCGT CGGGTCGGTC
TTTGTCAACA ATGAGGCCGG CAGCGGCGGC GCATTCGGCG GAATTGCGAC CGGCTTGCAG
GTCTACAACT CGCGGTTCAC GAACAACCGC GCCGCTGACG CGACGACGGG CGGCATTGTG
CGAGGGCATG GCGGCGCCAT TCATCTCGAT GGCGTCAGCA ATAGCTTCAA TCCGATCACC
GGTAACACTG TCGAGGTCTG CGGCAGCGTG TTCGATGGAA ACACTGCAAC GCGCGGCGGC
GGCGCACTCA AAGTGACGAT CTCCGATAAC CTGAACACGA AAGCGACCTA TGCCCGCTCG
ACCTTTAGCA ACAACCGGGT GCTCGACTCA CCGCCCGCTG AAGGTCACGG CGGCGCTATT
TATCATATCG AAGACGACCT GGCAGGTGGG GTGAATGAGG ACAATATCGA GATCCGCGAG
TCGACGTTCA ACGGCAACTA TGCCGCAAGG CAAGGAGGTG GCGTCTGGAT GTCGGTGAAC
GGTCGCGGAC GGGTGGTGAA CTCAACGTTC TTCGACAATC GCGCCGCAAC CGCCGGAACC
AACGTCATCG GTCAGGGAGG CGGCATGATC ATCAGCAGAG GGACCATCGA CATTATTCAC
GCGACGTTTG CGAGCAATTT CGCCACGTTT CAGGGTGGTG CGATCTTTGC CGGTAATCCG
TCAGCAGCGG CAGTGACACT GACCAGTTCA ATCTTCTCTG CTAATCGCCT CGATCCAACG
CACACGAACC CGGTCACAAC CGAGTTTCAG GGGTACCACA CGAACCGTGC ATTGCTCAAT
GGCGGCGGTA ACATTCAGTT TCCACGCACC AAAGCGCCCG ATTTCAATAA CGACATCAAC
AACCTGATAA CCTCACCGGC TTCGGCAATC CTCTTCCAGG ACCCACAACT CGCGCCGCTG
GCAAACAATG GCGGTCCCAC ACCGACGATG GCAATTGCGG CGTCCAGTCC GGGGTTCAAC
CGCGCTGCGG CGGCGACGTG CCCGGCAGCC GATCAGCGCG GCGTCATTCG TCCGCAGGGA
GGCGCCTGTG ATGTGGGAGC GTATGAACTG GTGTTGGCGC TATCGCTCAC GCCGCCGTTT
GTTGGCGTCG GCGAAGCAGG AACGGTTGTG ATCGTCTCCG GCGCCGGCTT CGACGCGAGC
AGCACGATCG TCATCGGCGG CGTTGAACGT CCGACAACGT TTGTCAGCGC CACTGAACTG
CGCACCACGC TCACTGCGGC TGATGTGGCG AACGCTGGCG ATCTGGAGGT GCGGGTCAGC
AACTCGGCGC TGCCGCCGGT CACGCTGCGC GTCCTGGCGC AGGTCTATCG TGGCTATTTG
CCGATCGTTC GTCGCTGA
 
Protein sequence
MNTSRPPSRW RRLLAHICLT ILCVSATVPP LPGMASAPAG APGAPPAACT PPIVPVTLVN 
PTVITSCTQA NLQAALATGG HITFDCGPHP VTIPITSPLV TSATRDIVLD GKGLITLDGG
GVTRILEKPF TPGSHIDKTS GNDLVIQNMR FINGRAPAAT KTQDDKARGG ALWVTSPGTR
LHIINSIFEN NRTTSMTDED NQGGAIYAGN IYETVIVGSV FVNNEAGSGG AFGGIATGLQ
VYNSRFTNNR AADATTGGIV RGHGGAIHLD GVSNSFNPIT GNTVEVCGSV FDGNTATRGG
GALKVTISDN LNTKATYARS TFSNNRVLDS PPAEGHGGAI YHIEDDLAGG VNEDNIEIRE
STFNGNYAAR QGGGVWMSVN GRGRVVNSTF FDNRAATAGT NVIGQGGGMI ISRGTIDIIH
ATFASNFATF QGGAIFAGNP SAAAVTLTSS IFSANRLDPT HTNPVTTEFQ GYHTNRALLN
GGGNIQFPRT KAPDFNNDIN NLITSPASAI LFQDPQLAPL ANNGGPTPTM AIAASSPGFN
RAAAATCPAA DQRGVIRPQG GACDVGAYEL VLALSLTPPF VGVGEAGTVV IVSGAGFDAS
STIVIGGVER PTTFVSATEL RTTLTAADVA NAGDLEVRVS NSALPPVTLR VLAQVYRGYL
PIVRR