Skip to navigationSkip to contentSkip to footer
Need Help?770-729-2992
MY ACCOUNT
CART
RayBiotech
RayBiotech
  • Products
    • Multiplex Assays
    • ELISA Kits
    • Proteins
    • Antibodies
    • Flow Cytometry
    • Assay Kits
    • Molecular Biology
    • Other Products
    • Human Cytokine Array C5
    • Human IL-6 ELISA
    • Human Inflammation Array Q3
    • Mouse Cytokine Array C3
    • Recombinant SARS-CoV-2 Spike Protein, S1 Subunit
    • Human Phosphorylation Pathway Profiling Array C55
    • Antibody Arrays
    • Proteome Profiling Arrays
    • Cytometric Bead Arrays
    • Epitope Mapping Peptide Arrays
    • Protein Arrays
    • PTM Multiplex Assays
    • NexaTag® Ultrasensitive Multiplex ELISA
    • Sandwich ELISA
    • Competitive ELISA
    • PTM ELISA
    • NexaTag® Ultrasensitive qIPCR ELISA
    • Indirect ELISA
    • Bridging ELISA
    • Cell-Based ELISA
    • Recombinant Proteins
    • GMP Proteins
    • Native Proteins
    • Peptides
    • Primary Antibodies
    • Secondary Antibodies
    • Flow Cytometry Antibodies
    • Isotype Controls
    • Antibody Pairs
    • Flow Cytometry Antibodies
    • Flow Cytometry Reagents
    • Flow Cytometry Assays
    • Cytometric Bead Arrays
    • Fluorescent Beads
    • Quantum Dots
    • Metabolism Assays
    • Ligand Binding Assays
    • Drug Antibody (DA) & ADA Assays
    • Transcription Factor Activity Assays
    • PCR Assays
    • Nucleic Acid Assays
    • Epigenetics
    • Mitochondrial Assays
    • Reagents
    • Cell Culture
    • Biospecimen Samples
    • Sample Collection & Preparation
    • Lysates
    • Mass Spectometry Reagents
    • Magnetic Beads
    • Lab Equipment and Supplies
    • Biological Buffers

    Featured Products

    Human Cytokine Array C5Human IL-6 ELISAHuman Inflammation Array Q3Mouse Cytokine Array C3Recombinant SARS-CoV-2 Spike Protein, S1 SubunitHuman Phosphorylation Pathway Profiling Array C55
  • Services
    • Multiplex Assay Services
    • ELISA Services
    • Protein Services
    • Antibody Services
    • CRO Services
    • Flow Cytometry Services
    • SIMOA – Single Molecule Array
    • Other Services
    • Quantitative Proteomics Services
    • Discovery Proteomics Services
    • Custom Arrays
    • Array Scanning and Analysis
    • Custom Protein Production
    • GMP Protein Production
    • IHC Controls
    • Stable Cell Line Development
    • Molecular Biology Services
    • Antibody Production
    • Recombinant Antibody Expression
    • Bulk Antibody Production
    • Antibody Conjugation Services
    • Cell Biology Services
    • Epitope Mapping Services
    • Biomarker Discovery
    • Cell Biology Services
    • Antibody Drug Development & Characterization
    • Pharmacokinetic and Pharmacodynamic Analysis
    • Assay Development Services
    • Molecular Biology Services
    • Multiomics Services
    • Diagnostic Assay Development
    • GMP Protein Production
    • Biostatistics and Bioinformatics
    • Bulk Antibody Production
    • QC Testing Services
    • Auto-Western Blot Service
    • COVID-19 Pseudovirus Service
    • Biospecimen Samples & Sample Services
    • RNA & DNA Methylation Detection Services
  • Research Areas
    • Cell Signaling Pathways
    • Areas of Interest
    • Akt Signaling Pathway
    • AMPK Signaling pathway
    • ErbB Signaling Pathway
    • Hedgehog Signaling Pathway
    • HIF1-alpha Signaling Pathway
    • IGFR1 Signaling Pathway
    • JAK-STAT Signaling Pathway
    • MAPK Signaling Pathway
    • mTOR Signaling Pathway
    • NF-kappaB Signaling Pathway
    • Notch Signaling Pathway
    • p53 Signaling Pathway
    • PKC Pathway
    • TGF-beta Signaling Pathway
    • Wnt/beta-Catenin Pathway
    • Antibody Drug Development
    • Autoimmunity & Inflammation
    • Cancer
    • Cardiovascular Disease
    • Infectious Disease & Vaccines
    • Neuroscience
    • Obesity
    • Post Translational Modifications
    • RNA & DNA Modifications
  • Resources
    • Assay Picker Tool
    • Sample Shipment Instructions
    • Sample Preparation Tips
    • Publications / Citations
    • Promotions
    • Resource Library
    • Learning Center
    • Manual Protocols
    • Array Analysis Tools
  • Contact Us
    • Distributors & Service Providers
  • About Us
    • Quality Systems
    • Business Partnerships
    • Careers
    • News & Events
Products
  • Multiplex Assays
  • ELISA Kits
  • Proteins and Peptides
  • Antibodies
  • Flow Cytometry
  • Assay Kits
  • Molecular Biology

Sign up to get promotions on your favorite research tools and research updates.

Sign up to Newsletter (opens in new tab)
Services
  • Multiplex Assay Services
  • ELISA Services
  • Antibody Services
  • Custom Protein Services
  • CRO Services
  • Flow Cytometry Services
  • Simoa – Single Molecule Array Services
Resources
  • Manual Protocols
  • Resource Library
  • Learning Center
  • Array Analysis Tools
  • Publications / Citations
  • Publications by RayBiotech Scientists
About Us
  • About RayBiotech
  • Careers
  • Quality Systems
  • Business Partnerships
  • News & Events
  • Reducing Our Emissions
Contact Us
  • Contacts
  • Distributors
  • Certified Service Providers
ISO 13485:2016cGMP
ISO 17025:2017CLIA
©2007-2026 RayBiotech, Inc. All rights reserved.Life Science Web Design By Supreme
  • Privacy Policy
  • Terms & Conditions
  • ISO Certification
  • Risk Free Guarantee
  • Promotions

Your cart is empty.

  • Home
  • Learning Center
  • Biostatistics & Bioinformatics
  • Random Forest Model
November 21, 2018|Biostatistics & Bioinformatics|Valerie Jones, PhD

Random Forest Model

What is it? Random forest consists of hundreds or more decision trees, with each tree using a random subset of data. All of the decision trees cast a vote on the classification of a sample; the majority vote wins.

When is it used? This analysis is one of the most commonly used models. It is used when 1) there are a lot of variables to consider (e.g., expression of thousands of proteins), 2) you only have moderate computing capacities, 3) you don’t want to analyze a separate set of samples for cross-validation, and 4) the groups are or are not normally distributed.

How does it work?

Random forest: Example

We analyze the protein profile of 1,000 proteins of 100 healthy patients and 100 cancer patients using an antibody-based microarray. We want to find biomarkers that will predict which future patients are healthy or diseased.

  1. Create a table where each row represents a protein and each column represents a patient.
  2. Assign groups. Here, you tell the software which samples are healthy and which have cancer.
  3. Center data by subtracting the mean of each patient dataset from itself. Now all datasets have a mean of 0.
  4. Scale data by dividing each patient dataset with its standard deviation. Now all datasets have a standard deviation of 1.
  5. Set aside 1/3 of the data. These samples will be used in Step 7 for cross-validation.
  6. Create decision trees using a subset of samples and variables at a time (Figure 1). The modeler determines the number of trees. Each tree will be created using a different random subset of data; the same sample can be chosen more than once to create 1 tree.
  7. Determine accuracy of the random forest. The samples set aside in Step 5 are tested against all of the decision trees. The accuracy of the random forest is the proportion of patients that were correctly identified by the random forest.

Leave a Reply

  • Apply Random Forest to samples with unknown health condition. There will be some trees that classify the patient as healthy, while other trees will classify the patient as diseased. The majority decision wins.
  • Random Forest model diagram

    What does the data look like? The Random Forest model is an ensemble of hundreds of trees that cannot be represented easily; however, the biomarkers used to create the model can be extracted during cross-validation and evaluation of the model's performance (e.g., via ROC curve analysis).