← All lessons

Lesson 5 of 8

Bayesian Methods in Clinical Research and Evidence Synthesis

Learning objectives

By the end of this lesson, you should be able to:

  • Explain how priors and accumulating data can be used in an adaptive clinical trial.
  • Describe how Bayesian meta-analysis represents heterogeneity and uncertainty.
  • Interpret posterior probabilities, credible intervals, and predictive intervals as answers to distinct research questions.
  • Distinguish probabilistic estimation from causal identification.
  • Identify assumptions that should be reported when Bayesian methods are used in evidence synthesis.

Estimated time: 35-45 minutes

Prerequisite: Lesson 4

Key terms: prior distribution, adaptive trial, posterior distribution, hierarchical model, heterogeneity, evidence synthesis

Why Bayesian Methods Matter in Clinical Research

Bayesian reasoning isn’t just a tool for better clinical judgment—it offers a powerful framework for improving how research is designed, analyzed, and synthesized. Frequentist and Bayesian approaches can both support rich estimation and adaptive designs, but they express uncertainty and answer inferential questions differently.

Bayesian methods make prior assumptions explicit and update probability distributions as new data emerge. They can shift the question from whether a result crosses a significance threshold to the probability that an effect exists, exceeds a clinically meaningful threshold, or will recur in a new setting. This can improve clinical relevance when the model and prior are defensible.

In this lesson, we explore how Bayesian thinking transforms:

Bayesian Clinical Trials: A More Adaptive Approach

Clinical trials can use fixed or adaptive designs under either frequentist or Bayesian frameworks. Bayesian methods are especially well suited to adaptive designs because posterior probabilities can be updated as prespecified interim data accumulate. Adaptations such as changing allocation, stopping early, or refining enrollment remain design choices that require explicit rules and calibration; they are not automatic consequences of using Bayesian statistics.

Key Advantages of Bayesian Trials:

Example: Adaptive Trial for Pain Management

Imagine a Bayesian trial for a new manual therapy technique. The study might:

  1. Begin with prior data from pilot studies and expert opinion.

  2. Continuously refine the estimated effect size as patients report outcomes.

  3. Adapt recruitment—perhaps oversampling subgroups who appear to respond best.

  4. Stop early if strong evidence accumulates for or against the treatment.

The potential result is a more efficient or ethically responsive trial, but only when the adaptation rules, prior assumptions, error characteristics, and decision thresholds are prospectively specified and evaluated.

Bayesian Meta-Analysis: Moving Beyond Pooled Averages

Both frequentist and Bayesian meta-analyses can weight studies by precision, estimate pooled effects, and model between-study heterogeneity. Their inferential interpretations differ: frequentist intervals describe procedure performance under repeated sampling, whereas Bayesian models produce posterior distributions conditional on the specified likelihood, prior, and model.

Bayesian meta-analyses take a different approach. They synthesize findings probabilistically, formally weighting prior knowledge, incorporating methodological nuance, and modeling heterogeneity more directly.

What Bayesian Meta-Analysis Can Add:

Example: Chronic Pain and Manual Therapy

A Bayesian meta-analysis on manual therapy might:

This produces a richer synthesis—focused not just on “if it works,” but how likely it is to work in real-world contexts.

Interactive activity

Bayesian Evidence Synthesis Lab

Combine a prior distribution with three study estimates, then distinguish uncertainty about the pooled mean from uncertainty in a new setting.

This is a simplified normal-normal random-effects model for instruction. Effect estimates are on an illustrative continuous scale where positive values favor the intervention. The model does not evaluate bias, study quality, publication bias, or causal identification.

Prior for the pooled mean effect

A smaller standard deviation makes the prior more concentrated and gives it more influence. A larger value makes it weaker.

Model and decision assumptions

Here, τ is supplied rather than estimated. The threshold is a decision assumption, not a value discovered by the model.

StudyEffect estimateStandard errorPrecision contribution
Study 1 29.5%
Study 2 48.4%
Study 3 19.5%

Prior precision contribution: 2.7%

Pooled posterior mean
+2.84
95% credible interval
+1.56 to +4.13
P(pooled mean > 0)
100.0%
P(pooled mean > +2.00)
90.1%
95% predictive interval
+1.23 to +4.46
P(new-setting effect > 0)
100.0%

Given this prior, these estimates, and τ = 0.5, the posterior probability that the pooled mean is greater than zero is 100.0%. The probability that it exceeds +2.00 is 90.1%.

Interrogate the synthesis

  • Make the prior skeptical and concentrated. How many precise, consistent studies does it take to move the posterior?
  • Load conflicting studies and increase τ. Why can the pooled mean remain fairly precise while prediction in a new setting becomes uncertain?
  • Compare P(effect > 0) with P(effect > MCID). Which question is more useful than a binary rejection decision?

Model: yᵢ ∼ Normal(μ, SEᵢ² + τ²), with a Normal prior on μ. The displayed “precision contributions” are each source’s share of total posterior precision. A full Bayesian random-effects analysis would generally assign a prior to τ and estimate it rather than setting it with a slider.

Bayesian Approaches to Causal Inference

Cause-and-effect reasoning is central to research, yet traditional statistical models often blur the line between association and causation.

Bayesian networks represent conditional dependencies within a specified graph and can update probabilities as evidence is entered. A network supports causal interpretation only when its arrows encode defensible causal assumptions and the relevant identification conditions are satisfied. Bayesian estimation does not turn an associational graph into evidence of causation.

Estimation Frameworks and Causal Models

Causal assumptions come from the study design and causal model, not from choosing a frequentist or Bayesian estimator. Either framework can estimate quantities defined by a defensible causal model:

Example: Stroke Rehabilitation

In a Bayesian causal model of motor recovery:

This aligns closely with real-world clinical inquiry—where causality is proposed, not proven, and must be refined iteratively.

Bayesian Reasoning in Critical Realist Reviews (CCRRs)

Critical realist reviews aim to move beyond surface-level correlations to uncover why interventions work, for whom, and under what conditions. They emphasize causal mechanisms, context, and structural influences—making Bayesian logic a natural fit.

Why Bayesian Reasoning Supports CCRRs:

Example: Manual Therapy Revisited

A frequentist review might conclude that manual therapy has mixed evidence for low back pain. A Bayesian CCRR would instead:

  1. Begin with priors informed by biomechanical, neurophysiological, and contextual mechanisms.

  2. Adjust estimates based on study context (e.g., patient type, clinician skill, setting).

  3. Identify plausible generative mechanisms in specific populations, even where pooled effects are “non-significant.”

The result? A deeper understanding of why and when manual therapy might work—not just whether it does. Meaning - research evidence doesn’t just have to answer “does it work” but it should also dig into why, when, how, where and to what extent.

From Bayesian Research to Models4PT

Models4PT is an open research platform for building, curating, integrating, and maintaining computable population-level causal knowledge in physical therapy and rehabilitation science. Its canonical product is curated knowledge, not a collection of diagrams or an automated clinical decision tool. Models4PT keeps concepts, measurements, mechanisms, causal claims, evidence, provenance, uncertainty, disagreement, and researcher review connected.

Models4PT is currently in an early research and software-design stage, with a tested initial domain model rather than a deployable platform. It constructs population-level knowledge; patient-specific diagnosis, prognosis, treatment recommendations, and probabilistic reasoning belong to the separate Clinical Inference Engine.

What’s Next?

Bayesian methods can update uncertainty, but they do not determine whether the variables in a model are the right ones or whether a population result applies to a particular context. Lesson 6 turns to mechanisms, context, and the limits of population models.