Blog
Research Methodology Insights
Practical guides on causal inference, study design, and statistical methods — written for researchers who care about getting the design right before touching the model.
The Hazard-Ratio Autopsy: When Delayed Effects Make One Number Misleading
A practical trial autopsy for non-proportional hazards. Read the curves, risk sets, landmarks, and restricted mean survival time before translating one hazard ratio into a constant treatment effect.
Latest picks
Freshest in the archiveHierarchical Testing in Clinical Trials: When a Significant Secondary Endpoint Is Still Descriptive
A practical hierarchical testing guide for clinical researchers. Reconstruct the prespecified testing path before treating a small p-value on a secondary endpoint as confirmatory evidence.
Randomized Withdrawal Trials: Why a Relapse-Prevention Win Is Not a New-Patient Effect
A practical randomized withdrawal trial guide for clinical researchers. Audit the run-in, responder enrichment, withdrawal contrast, safety, and target population before generalizing a maintenance-effect claim.
Allocation Concealment: The Randomized-Trial Safeguard That Works Before Assignment
A practical allocation concealment guide for clinical researchers and peer reviewers. Separate sequence generation, concealment, implementation, and blinding before trusting the word randomized.
Browse all guides
Filter by method family or search for the exact problem you're trying to solve.
The Denominator Illusion: Why 20,000 Measurements May Still Mean 200 Patients
A practical unit-of-analysis guide for clinical researchers. Separate rows, patients, clusters, and target populations before repeated observations create false precision.
Ecological Fallacy: When Hospital-Level Data Become Patient-Level Advice
A practical ecological fallacy guide for clinical researchers. Match the unit of analysis, exposure, outcome, and claim before turning group-level associations into patient-level advice.
Delayed Entry in Survival Analysis: Why Nobody Is at Risk Before They Enter
A practical guide to left truncation in clinical survival analysis. Align the origin, entry time, event time, and risk set before interpreting Kaplan–Meier curves or Cox models from prevalent cohorts.
Cluster-Trial Recruitment Bias: When the Clinic Knows the Assignment Before the Patient Enters
A practical guide to post-randomization recruitment bias in cluster trials. Audit timing, allocation awareness, eligibility, consent, patient denominators, and the limits of statistical adjustment.
Target Trial Emulation Cannot Randomize Clinical Judgment: A Pertussis Study Audit
A practical target-trial emulation audit using an infant pertussis study. Check clinical-judgment confounding, propensity-score overlap, endpoint timing, sparse outcomes, and claim strength.
Vaccine Effectiveness Without Matching: Why Calendar Time Comes Before Pairing
A practical guide to calendar time in vaccine-effectiveness studies. Learn how changing uptake and infection hazards alter risk sets, estimands, and target-trial conclusions.
Modified Intention-to-Treat: When Randomization Starts Losing Patients After the Fact
A practical guide to modified intention-to-treat analyses for clinical researchers. Learn how post-randomization exclusions weaken trial credibility, when exclusions may be defensible, and what reviewers should demand before trusting a modified analysis population.
Crossover Trials: When Every Patient Is Their Own Control—and Their Own Carryover Problem
A practical guide to crossover trials for clinical researchers. Audit treatment reversibility, washout, carryover, period effects, sequence, dropout, and paired analysis before trusting an efficient within-patient comparison.
Estimands in Meta-Analysis: When a Shared PICO Still Pools Different Questions
A practical guide to estimands in meta-analysis. Learn how treatment-policy and hypothetical strategies can make trials with the same PICO answer different questions, and how to audit the pool before combining effects.
Causal Identification: Why Longitudinal Data Still Need a Design
A practical causal identification guide for longitudinal clinical research. Learn why measurement order does not create comparability, what repeated observations add, and how to audit causal claims with five reviewer questions.
Specification Curve Analysis: When One Model Hides a Vibration of Effects
A practical guide to specification curve analysis in clinical research. Learn how to define reasonable analyses, separate robustness checks from estimand changes, and audit results that move across defensible choices.
AI Before–After Studies: When Faster Care Is Not Yet an AI Effect
A practical guide to evaluating healthcare AI after deployment. Audit pre-trends, concurrent comparisons, co-interventions, outcome measurement, and the claim ceiling of before–after and interrupted time-series designs.
Comparator Selection in Observational Studies: Why the Control Group Changes the Question
A practical guide to comparator selection in observational studies. Learn how active comparators change the estimand, confounding structure, and interpretation of real-world evidence.
AI Literature Search: Why Plausible Citations Are Not Evidence Coverage
A practical guide to auditing AI literature search in clinical research. Separate valid citations from evidence coverage, measure retrieval recall, and demand a reproducible search trail.
Recurrent Events in Clinical Trials: When Time to First Event Hides Disease Burden
A practical guide to recurrent-event analysis in clinical trials. Learn when time to first event discards patient burden, how death changes the estimand, and what reviewers should demand.
AI Surveillance Models: Why a High AUC Cannot Justify Fewer Follow-Up Visits
A practical guide to evaluating AI surveillance models. Learn why high AUC is not enough to reduce follow-up, and audit calibration, thresholds, missed failures, utility, and prospective impact.
Treatment History Is Part of the Estimand: Why Starting and Switching Answer Different Questions
A practical guide to treatment history in target trial emulation. Learn why initiators and switchers define different estimands, how sequential trials align switching decisions, and why relative and absolute risk must travel together.
Time-Varying Treatment Dose: When Recovery Gets Counted as the Intervention
A practical guide to time-varying treatment dose in clinical research. Learn why accumulated exposure can encode recovery, how baseline matching misses treatment-confounder feedback, and what reviewers should demand from longitudinal dose-response claims.
Diagnostic Cutoff Selection: When the Same Data Chooses and Grades the Threshold
A practical guide to diagnostic cutoff selection. Learn why searching and grading a threshold in the same sample inflates apparent performance, why Youden index is not clinical utility, and what reviewers should demand before implementation.
Test-Negative Designs: When the Control Group Is Chosen by the Testing Door
A practical guide to auditing test-negative vaccine-effectiveness studies. Check the clinical case definition, testing pathway, control illness, calendar time, and estimand before treating 1 minus an odds ratio as protection.
Treatment-Timing Effects: When “Earlier Is Better” Needs a Fair Clock
A practical guide to auditing earlier-is-better claims in observational studies. Test time zero, treatment strategies, evolving clinical decisions, positivity, and timing curves before reading a treatment gradient as biological.
Broad Benefit Patterns in Observational Studies: When One Treatment Appears to Prevent Everything
A practical guide to auditing outcome-wide benefit patterns in real-world evidence. Learn what clean negative controls can and cannot prove, how shared bias can move many endpoints together, and when a broad pattern should remain hypothesis-generating.
Discordant Clinical Endpoints: When One Study Produces Two Treatment Winners
A practical guide to discordant clinical endpoints using a matched MS registry study. Audit endpoint hierarchy, estimand alignment, ascertainment, and absolute effects before turning a mixed outcome dashboard into one treatment winner.
How to Audit a Target-Trial Emulation: Four Decisions Behind One Mortality Estimate
A practical target-trial emulation audit using a dialysis mortality study. Translate a day-91 landmark, sustained-treatment threshold, competing event, and center context into the question actually estimated.
Co-Intervention Bias: When the Treatment Arm Gets More Than the Treatment
A practical co-intervention bias guide for clinical trial reviewers. Learn when unequal concomitant care distorts a treatment claim, belongs to the strategy, or changes the estimand.
Separation in Logistic Regression: When Perfect Prediction Breaks the Odds Ratio
A practical guide to separation in logistic regression for clinical researchers. Learn why sparse or perfectly predictive data produce infinite estimates, how to diagnose the failure, and what reviewers should demand before trusting a dramatic odds ratio.
Interval Censoring: When the Visit Date Pretends to Be the Event Date
A practical interval censoring guide for clinical researchers. Learn why periodic assessments create event-time windows, how naive date substitution manufactures precision, and what reviewers should demand before trusting time-to-event claims.
Multi-State Models: When One Time-to-Event Endpoint Hides the Clinical Path
A practical multi-state models guide for clinical researchers. Learn how state occupation, transition hazards, and time in state reveal clinical paths hidden by one survival endpoint.
Endpoint Adjudication: When a Blinded Committee Cannot Rescue Biased Event Capture
A practical endpoint adjudication guide for clinical researchers. Learn why blinded central review cannot repair unequal event capture, incomplete evidence dossiers, or post hoc endpoint rules.
Desirability of Outcome Ranking: When Benefit–Risk Depends on Who Ranks the Outcomes
A practical guide to desirability of outcome ranking (DOOR) in clinical trials. Learn how whole-patient outcome ranks encode benefit–risk judgments, how to interpret DOOR probability, and what reviewers should demand before trusting one summary number.
Risk-Set Matching: Why the Comparison Group Must Follow the Clock
A practical risk-set matching guide for time-varying treatment. Learn why later-treated patients can be valid comparators now, how fixed exposure groups create timing bias, and what reviewers should demand.
CONSORT 2025 for Reviewers: A Practical Checklist for Interpretable Trials
A practical CONSORT 2025 reviewer checklist for randomized trials. Find the reporting gaps that change what you can infer about allocation, treatment, outcomes, missing data, and analysis.
Before the Model: A Five-Gate Study Design Audit for Clinical Research
A practical five-gate study design audit for clinical researchers. Check the question, population, comparison, measurement, and follow-up before a sophisticated model gives a fragile design false confidence.
Measurement Invariance: When the Same Clinical Score Means Different Things
A practical guide to measurement invariance in clinical research. Learn why the same score may not be comparable across groups, sites, languages, or devices, and what reviewers should demand before trusting the comparison.
Causal Readiness: When a Huge Linked Dataset Still Cannot Identify an Effect
A practical guide to causal readiness in linked health and administrative data. Learn why scale and propensity-score overlap are not enough when treatment, need, comparators, or outcomes are poorly measured.
When More Covariates Break Positivity: Representation-Induced Overlap Failure in Clinical Text
A practical guide to representation-induced positivity failure in clinical text. Learn why richer embeddings can encode treatment, shrink common support, and make a causal adjustment less trustworthy.
Healthy Screenee Bias: When Screening Attendance Looks Like Screening Benefit
A practical guide to healthy screenee bias for clinical researchers. Learn why people who attend screening can look healthier before the screen is credited, how this differs from lead-time bias, and what reviewers should demand from observational screening studies.
Predicted Treatment Benefit: When a Risk Model Is Not a Treatment Recommendation
A practical guide to separating predicted outcome risk from predicted treatment benefit. Learn why a high-risk patient is not automatically a high-benefit patient, how risk modeling and effect modeling differ, and what reviewers should demand before trusting a personalized treatment claim.
Additive Interaction: When “No Interaction” Depends on the Scale
A practical guide to additive and multiplicative interaction in clinical research. Learn why a null product term can hide clinically important effect modification, how to read RERI, and what reviewers should demand before trusting a joint-exposure claim.
Small Number of Clusters: When 500 Patients Still Behave Like 10 Sites
A practical guide to inference with a small number of clusters in clinical research. Learn why patients inside the same site are not independent evidence, why default cluster-robust standard errors can be too optimistic, and what reviewers should demand before trusting a cluster-level result.
Heterogeneous Treatment Effects: Validate the Average Effect Before Trusting the Subgroups
A practical guide to validating heterogeneous treatment-effect claims in target trial emulation. Learn why a failed average-effect benchmark should stop a subgroup story, how subgroup effects must reconcile with their parent estimate, and what reviewers should demand before trusting personalized treatment claims.
Estimation Over Testing: Four Habits That Quietly Break Clinical Statistics
A practical critique of four entrenched habits in clinical biostatistics — significance testing where the real question is estimation, prediction accuracy scored far from the bedside, meta-analysis on autopilot, and the false precision of point estimates that assume away what cannot be tested — with an interactive explorer that reads one trial three ways and what to do instead.
Covariate Balance Diagnostics: When the Love Plot Says Balanced but the Groups Still Differ
A practical guide to covariate balance diagnostics after propensity-score matching or weighting. Covers why a Love plot of standardized mean differences can read balanced while distributions differ in spread and tails, why you should check variance ratios and distributional distances and interactions, why you should never significance-test balance because the p-value tracks sample size, and what reviewers should demand.
The Will Rogers Phenomenon: When Better Staging Improves Every Group and Nobody Lives Longer
A practical guide to the Will Rogers phenomenon (stage migration) for clinical researchers. Covers why sharper classification can lift survival in every stage while overall survival stays flat, how to separate stage migration from real progress, how it differs from lead-time bias and overdiagnosis, and what reviewers should demand when outcomes are compared across eras or cohorts.
The Imperfect Gold Standard: Measuring a Test When the Reference Is Wrong Too
A practical guide to imperfect reference standards and latent class analysis for clinical researchers. Covers why apparent sensitivity and specificity are biased when the gold standard is itself flawed, why conditional dependence between tests flips the bias from pessimistic to optimistic, when latent class analysis helps and when it fails, and what reviewers should demand.
Incorporation Bias: When a Test Helps Write Its Own Answer Key
A practical guide to incorporation bias for clinical researchers. Covers why sensitivity and specificity both inflate toward 100% when the index test is used to define the reference standard, how it differs from verification and spectrum bias, why there is no clean statistical correction, and what reviewers should demand before trusting a diagnostic accuracy.
Verification Bias: When the Test Under Study Decides Who Gets the Gold Standard
A practical guide to verification bias (workup bias) for clinical researchers. Covers why sensitivity is inflated and specificity deflated when the index test drives who gets the reference standard, the Begg-Greenes correction, differential verification, and what reviewers should demand before trusting a diagnostic accuracy.
Spectrum Bias: Why a Test’s Accuracy Is Not a Property of the Test
A practical guide to spectrum bias for clinical researchers. Covers why sensitivity and specificity shift with the case-mix of who was enrolled, the two-gate case-control trap, how curated data inflates AI-diagnostic performance, and what reviewers should demand before trusting a reported accuracy.
Simpson's Paradox: When Every Subgroup Says One Thing and the Total Says the Opposite
A practical guide to Simpson's paradox for clinical researchers. Covers why a treatment can help in every subgroup yet look harmful pooled, why 'always stratify' is wrong, and how the causal role of the stratifier — confounder, mediator, or collider — decides which table to trust.
The Table 2 Fallacy: When Every Adjusted Coefficient Looks Like a Cause
A practical guide to the Table 2 fallacy for clinical researchers. Covers why secondary coefficients in an adjusted model are not causal effects, mutual adjustment, mediator and confounder mismatch, and what reviewers should demand before reading a covariate row as a finding.
Bayesian Borrowing: When Historical Data Starts Spending Credibility It Did Not Earn
A practical guide to Bayesian borrowing for clinical researchers. Covers exchangeability, commensurate priors, historical controls, calendar-time drift, and what reviewers should demand before trusting extra certainty borrowed from earlier data.
PROBAST: When a Prediction Model Paper Looks Ready Before It Earns Trust
A practical guide to PROBAST for clinical researchers. Covers participant selection, predictor leakage, outcome definition, overfitting, calibration, and what reviewers should demand before trusting a clinical prediction model.
Informative Cluster Size: When the Biggest Sites Start Writing the Result
A practical guide to informative cluster size for clinical researchers. Covers why larger centers can quietly dominate treatment effects, how weighting changes the estimand, and what reviewers should demand before trusting clustered results.
Baseline Adjustment in Randomized Trials: Why Change From Baseline Keeps Losing to ANCOVA
A practical guide to baseline adjustment in randomized trials. Covers ANCOVA versus change scores, percent change traps, responder thresholds, and what reviewers should demand before trusting a tidy efficacy claim.
Depletion of Susceptibles: When Early Harm Vanishes Because the Vulnerable Patients Are Already Gone
A practical guide to depletion of susceptibles for clinical researchers. Covers front-loaded harm, survivor selection, why later follow-up can falsely reassure, and what reviewers should demand before trusting a calming hazard curve.
Indirectness in Clinical Evidence: When a Good Study Answers the Wrong Question
A practical guide to indirectness in clinical evidence for clinical researchers. Covers PICO mismatch, outdated comparators, surrogate outcomes, and what reviewers should demand before trusting an applicable-sounding conclusion.
Complete-Case Analysis: When Missing Data Quietly Changes the Study Population
A practical guide to complete-case analysis for clinical researchers. Covers when dropping incomplete records changes the study population, how endpoint missingness becomes selection bias, and what reviewers should demand before trusting the estimate.
Platform Trials: When a Shared Control Stops Being the Same Comparison
A practical guide to platform trials for clinical researchers. Covers nonconcurrent controls, changing standard of care, case-mix drift, and what reviewers should demand before trusting an adaptive-trial headline.
Nonproportional Hazards: When One Hazard Ratio Pretends the Treatment Effect Never Changes
A practical guide to nonproportional hazards for clinical researchers. Covers delayed effects, crossing curves, waning benefit, why one hazard ratio can mislead, and what reviewers should demand instead.
Futility Stopping in Clinical Trials: When “No Signal Yet” Starts Pretending the Question Is Answered
A practical guide to futility stopping in clinical trials. Covers conditional power, delayed effects, optimistic design assumptions, and what reviewers should demand before trusting a trial stopped for futility.
Nested Case-Control Design: When Cheap Control Sampling Still Has to Respect Event Time
A practical guide to nested case-control design for clinical researchers. Covers incidence-density sampling, risk-set control selection, how it differs from case-cohort design, and what reviewers should demand before trusting the result.
Transitivity in Network Meta-Analysis: When Indirect Comparisons Pretend the Trials Were Exchangeable
A practical guide to transitivity in network meta-analysis for clinical researchers. Covers effect modifiers, shared comparators, indirect comparison failure modes, and what reviewers should demand before trusting rankings.
Benchmarking Target Trial Emulation: When a Trial Copy Never Checks It Can Reproduce the Known Answer
A practical guide to benchmarking target trial emulation for clinical researchers. Covers why benchmark replication matters, what it cannot prove, and what reviewers should demand before trusting an observational extension.
Weak Instruments and Physician Preference IVs: When Treatment Movement Is Not Yet Causal Credibility
A practical guide to weak instruments and physician-preference IVs for clinical researchers. Covers first-stage weakness, exclusion leakage, local interpretation, and what reviewers should demand before trusting an IV claim.
Noninferiority Margins: When “Not Much Worse” Starts Giving Away Too Much
A practical guide to noninferiority margins for clinical researchers. Covers margin justification, assay sensitivity, constancy, biocreep, and what reviewers should demand before trusting a noninferiority win.
Win Ratio: When a Hierarchical Composite Endpoint Sounds Harder Than It Really Is
A practical guide to win ratio for clinical researchers. Covers hierarchical composite endpoints, pairwise priorities, soft-tier distortion, and what reviewers should demand before trusting a prioritized endpoint headline.
Repeated Eligibility in Target Trial Emulation: When One Patient Quietly Enters the Trial Again
A practical guide to repeated eligibility in target trial emulation for clinical researchers. Covers re-entry rules, overlapping follow-up, carryover, and what reviewers should demand before trusting a sequential-trial analysis.
ROBINS-I: When an Observational Effect Estimate Is Too Biased to Grade Casually
A practical guide to ROBINS-I for clinical researchers. Covers the seven bias domains, why confounding and time zero usually dominate, how overall judgments are formed, and what reviewers should demand before trusting non-randomized evidence.
M-Bias: When “Adjusted for Severity” Is Actually the Problem
A practical guide to M-bias for clinical researchers. Covers collider structures, severity-score traps, selected cohorts, and what reviewers should demand before trusting an adjusted estimate.
Stepped-Wedge Cluster Trials: When Rollout Timing Starts Competing With the Intervention
A practical guide to stepped-wedge cluster trials for clinical researchers. Covers secular trends, rollout order, learning effects, contamination, and what reviewers should demand before trusting a tidy implementation-era benefit.
Calibration Drift: When a Good Model Keeps the Right Rank and Still Gives the Wrong Risk
A practical guide to calibration drift for clinical researchers. Covers baseline-risk shift, calibration slope failure, threshold consequences, and what reviewers should demand before trusting deployment-ready prediction claims.
Response-Adaptive Randomization: When a Trial Starts Chasing Its Early Winners
A practical guide to response-adaptive randomization for clinical researchers. Covers delayed outcomes, temporal drift, instability, ethical claims, and what reviewers should demand before trusting an adaptive allocation design.
Case-Cohort Design: When Measuring Everyone Is the Wrong Expense
A practical guide to case-cohort design for clinical researchers. Covers when a random subcohort is more honest than measuring everyone, how it differs from nested case-control sampling, and what reviewers should demand before trusting the result.
Overlap Weighting: When the Average Treatment Effect Stops Being the Honest Question
A practical guide to overlap weighting for clinical researchers. Covers why overlap weighting can stabilize observational comparisons, how it changes the target population, and what reviewers should demand before trusting a trimmed-looking causal claim.
Augmented Inverse Probability Weighting: When “Doubly Robust” Starts Hiding Which Model Failed
A practical guide to augmented inverse probability weighting for clinical researchers. Covers what AIPW estimates, why doubly robust does not mean low-risk, and what reviewers should demand before trusting the label.
Case-Time-Control Design: When Case-Crossover Starts Confusing Time Trends with Treatment Effects
A practical guide to the case-time-control design for clinical researchers. Covers exposure-time trends, referent sampling, protopathic bias, and what reviewers should demand before trusting a self-matched trigger analysis.
Triangulation in Clinical Research: When One Elegant Design Still Leaves the Same Blind Spot
A practical guide to triangulation in clinical research. Covers what counts as genuinely complementary evidence, how to map designs to specific threats, and what reviewers should demand before trusting “robustness” claims.
Guideline Recommendation Strength: When “Strongly Recommend” Starts Outrunning the Evidence
A practical guide to recommendation strength in clinical guidelines. Covers certainty of evidence, benefit-harm tradeoffs, patient values, implementation burden, and what reviewers should demand before trusting a forceful recommendation.
Differential Misclassification: When One Study Arm Gets More Chances to Be Wrong
A practical guide to differential misclassification for clinical researchers. Covers arm-specific outcome detection, adjudication asymmetry, false positives, missed events, and what reviewers should demand before trusting an effect estimate.
Podcast with Saud Alomairah: The Physician-Researcher in the Age of AI
A short note on Anas Alzahrani joining Saud Alomairah on Afiyah Tech for an Arabic conversation about medicine, research, artificial intelligence, and rigorous clinical thinking.
Adaptive Enrichment Trials: When Precision for One Subgroup Pretends to Be Evidence for Everyone
A practical guide to adaptive enrichment trials for clinical researchers. Covers predictive versus prognostic enrichment, assay timing, multiplicity, external validity, and what reviewers should demand before trusting a biomarker-selected win.
Treatment-Induced Mediator-Outcome Confounding: When Mediation Analysis Starts Chasing the Consequences of Treatment
A practical guide to treatment-induced mediator-outcome confounding for clinical researchers. Covers why natural direct and indirect effects fail when treatment changes later severity, toxicity, adherence, or surveillance that affect both the mediator and outcome.
Surrogate Endpoints: When a Biomarker Improvement Pretends to Be Patient Benefit
A practical guide to surrogate endpoints for clinical researchers. Covers validated versus merely plausible surrogates, classic failure modes, and what reviewers should demand before trusting a biomarker-driven trial claim.
Quantitative Bias Analysis: When “Residual Confounding” Needs a Number, Not a Shrug
A practical guide to quantitative bias analysis for clinical researchers. Covers simple bias factors, plausible unmeasured confounding scenarios, and what reviewers should demand before trusting a causal claim that survives only because nobody quantified the threat.
Data Leakage in Clinical Prediction Models: When the Model Learns the Future
A practical guide to data leakage in clinical prediction models for clinical researchers. Covers post-outcome features, workflow proxies, validation traps, and what reviewers should demand before trusting a headline AUC.
Net Reclassification Improvement: When a New Biomarker Wins by Moving Patients Between the Wrong Boxes
A practical guide to net reclassification improvement for clinical researchers. Covers event and non-event NRI, arbitrary risk categories, overtreatment traps, and what reviewers should demand before trusting claims that a new model improved classification.
AI-Assisted Methods Review: What LLMs Can Catch, What They Cannot, and Where Judgment Still Matters
A practical guide to AI-assisted methods review for clinical researchers. Covers where LLMs help with structural critique, where source verification and causal judgment still require humans, and what reviewers should demand before trusting AI-generated methodological comments.
Decision Curve Analysis: When a Better AUC Still Makes Worse Clinical Decisions
A practical guide to decision curve analysis for clinical researchers. Covers net benefit, threshold probability, when prediction models fail to beat treat-all or treat-none strategies, and what reviewers should demand before trusting claims of clinical utility.
Channeling Bias: When the Newer Treatment Inherits the Easier Patients
A practical guide to channeling bias for clinical researchers. Covers preferential prescribing, formulary-era drift, specialist selection, and what reviewers should demand before trusting observational comparisons of newer therapies.
When Death Changes the Question: Competing Risks, Intercurrent Events, and Truncation by Death
A practical guide to the boundary between competing risks, intercurrent events, and truncation by death for clinical researchers. Covers when death changes risk sets, when it makes later outcomes undefined, and what reviewers should demand instead of vague censoring language.
Jump-to-Reference Imputation: When Missing Outcomes Start Borrowing the Control Arm's Future
A practical guide to jump-to-reference imputation for clinical researchers. Covers what J2R assumes after treatment discontinuation, when it helps sensitivity analysis, and when it quietly answers the wrong estimand.
Fragility Index: When One or Two Events Carry More Confidence Than They Should
A practical guide to the fragility index for clinical researchers. Covers event-flip sensitivity, loss to follow-up, effect-size context, and what reviewers should demand before trusting a barely significant trial.
Consistency and Treatment Versioning: When One Exposure Label Hides Several Different Interventions
A practical guide to consistency and treatment versioning for clinical researchers. Covers when an exposure is too vaguely defined for causal interpretation, how hidden intervention versions distort transportability, and what reviewers should demand instead.
Multiple Testing in Clinical Trials: When One Positive Endpoint Is Just the Loudest Coin Flip
A practical guide to multiple testing in clinical trials for clinical researchers. Covers endpoint families, subgroup fishing, interim looks, alpha control, and what reviewers should demand before trusting a lone positive result.
Confounding by Contraindication: When the Untreated Group Is Too Fragile for the Therapy
A practical guide to confounding by contraindication for clinical researchers. Covers how treatment avoidance in high-risk patients can make therapies look safer or more effective than they are, and what reviewers should demand instead.
Intercurrent Events in Clinical Trials: When Rescue Therapy and Death Are Not Missing Data
A practical guide to intercurrent events for clinical researchers. Covers rescue therapy, treatment switching, discontinuation, death, estimand strategy choices, and what reviewers should demand before trusting the headline effect.
Last Observation Carried Forward: When Yesterday's Outcome Pretends the Patient Stopped Changing
A practical guide to last observation carried forward for clinical researchers. Covers why LOCF fails as missing-data strategy, how it can exaggerate or dilute treatment effects, and what reviewers should demand instead.
Time Zero Alignment: When Your Cohort Starts Counting Before Treatment Does
A practical guide to time zero alignment for clinical researchers. Covers eligibility, treatment assignment, delayed initiation, immortal time, and what reviewers should demand before trusting a real-world effect estimate.
Early Stopping for Benefit: When a Trial Quits While the Effect Is Still on Its Best Behavior
A practical guide to early stopping for benefit in clinical trials. Covers interim looks, alpha spending, exaggerated effect sizes, immature follow-up, and what reviewers should demand before trusting a triumphant stop.
Informative Visit Processes: When Who Shows Up Starts Writing the Results
A practical guide to informative visit processes for clinical researchers. Covers endogenous follow-up, unequal observation schedules, visit-triggered outcome capture, inverse-intensity thinking, and what reviewers should demand before trusting longitudinal real-world results.
External Control Arms: When a Comparison Group Arrives from Another Universe
A practical guide to external control arms for clinical researchers. Covers historical and real-world comparators, design drift, prognostic imbalance, endpoint mismatch, and what reviewers should demand before trusting single-arm success stories.
Stochastic Interventions: When “Treat Everyone” Is Not the Policy Question
A practical guide to stochastic interventions for clinical researchers. Covers when deterministic treatment rules become unrealistic, how probability-shift interventions preserve positivity, and what reviewers should demand before trusting policy-effect claims.
Missing Indicator Method: When an NA Flag Pretends to Be Missing-Data Strategy
A practical guide to the missing-indicator method for clinical researchers. Covers why NA flags fail for confounding control, when they leave residual bias, and what reviewers should demand before trusting a covariate-adjusted result.
Baseline Covariate Windows: When “Pre-Treatment” Variables Arrive Fashionably Late
A practical guide to baseline covariate windows for clinical researchers. Covers lookback periods, stale severity measures, same-day contamination, confounding capture, and what reviewers should demand before trusting “adjusted” observational results.
Treatment Switching in Oncology Trials: When Overall Survival Becomes a Rescue Protocol Audit
A practical guide to treatment switching in oncology trials for clinical researchers. Covers crossover, overall survival dilution, ITT versus hypothetical estimands, RPSFTM, IPCW, two-stage estimation, and what reviewers should demand before trusting an adjusted survival claim.
Run-In Periods: When Your Trial Randomizes the Easy Patients First
A practical guide to run-in periods for clinical researchers. Covers adherence enrichment, tolerability selection, estimand drift, external validity, and what reviewers should demand before trusting a polished randomized cohort.
Washout Periods: When “New Use” Is Just Old Use with Better PR
A practical guide to washout periods for clinical researchers. Covers new-user definitions, refill cycles, intermittent treatment, data-history limits, and what reviewers should demand before trusting an incident-user cohort.
Exposure Lagging: When Your Induction Window Becomes Wishful Thinking
A practical guide to exposure lagging for clinical researchers. Covers induction periods, reverse causation, protopathic bias, estimand drift, and what reviewers should demand before trusting a lagged analysis.
Responder Analyses: When a Cutoff Turns a Clinical Gradient into a Headline
A practical guide to responder analyses for clinical researchers. Covers dichotomizing continuous outcomes, post hoc thresholds, baseline dependence, power loss, and what reviewers should demand before trusting "X% achieved response" claims.
Healthy Adherer Bias: When Persistence Looks Like Pharmacology
A practical guide to healthy adherer bias for clinical researchers. Covers why adherent patients often look healthier before the treatment effect is even estimated, how this differs from confounding by indication, and what reviewers should demand before trusting adherence-based benefit claims.
Grace Periods in Target Trial Emulation: Clinical Realism or Future Information in Disguise?
A practical guide to grace periods in target trial emulation for clinical researchers. Covers when a grace window is defensible, when it becomes immortal time in formalwear, and what reviewers should demand before trusting the result.
Index Event Bias: When Your Cohort Already Selected the Wrong Comparison
A practical guide to index event bias for clinical researchers. Covers recurrence-risk paradoxes, conditioning on the first event, secondary prevention cohorts, and what reviewers should demand before trusting protective-looking associations inside diseased cohorts.
Calendar Time Confounding: When Secular Trends Pretend Your Intervention Worked
A practical guide to calendar time confounding for clinical researchers. Covers secular trends, treatment diffusion, concurrent comparators, and what reviewers should demand before trusting real-world benefit that may just reflect a later era.
Surveillance Bias: When One Group Gets More Chances to Become a Case
A practical guide to surveillance bias for clinical researchers. Covers differential testing, follow-up intensity, diagnosis-based outcomes, and what reviewers should demand before trusting higher event rates.
Outcome Switching: When the Primary Endpoint Moves After the Results Get Interesting
A practical guide to outcome switching for clinical researchers. Covers endpoint shopping, selective reporting, protocol drift, and what reviewers should demand before trusting a late-breaking primary outcome.
Overdiagnosis: When Finding More Disease Does Not Mean Saving More Lives
A practical guide to overdiagnosis for clinical researchers. Covers how screening can raise incidence and improve survival statistics without reducing mortality, how to separate lead-time from true overdiagnosis, and what reviewers should demand before trusting the headline.
Prevalent-User Bias: When Your Drug Study Starts After the Interesting Harm Already Happened
A practical guide to prevalent-user bias for clinical researchers. Covers depletion of susceptibles, survivor selection, post-treatment baseline covariates, and what reviewers should demand before trusting late-entry treatment cohorts.
Lead-Time Bias: When Earlier Diagnosis Pretends to Be Better Survival
A practical guide to lead-time bias for clinical researchers. Covers why screening can improve survival statistics without reducing mortality, how to separate earlier detection from real benefit, and what reviewers should demand before trusting the headline.
Clone-Censor-Weight: The Target Trial Fix That Still Breaks When You Use It Casually
A practical guide to clone-censor-weight for clinical researchers. Covers when the design is needed, how cloning and artificial censoring work, where immortal time bias reappears, and what reviewers should demand before trusting a target trial emulation.
MNAR Sensitivity Analysis: Because “We Assumed MAR” Is Not a Results Section
A practical guide to MNAR sensitivity analysis for clinical researchers. Covers when multiple imputation under MAR is not enough, how to think about missing not at random assumptions, and what reviewers should demand before trusting complete-case comfort.
Subgroup Analysis: When “Personalized” Findings Are Mostly Multiplicity Wearing a Stethoscope
A practical guide to subgroup analysis for clinical researchers. Covers interaction testing, multiplicity, power failure, post hoc storytelling, and what reviewers should demand before trusting treatment-effect heterogeneity claims.
Noncollapsibility of Odds Ratios: Why Adjustment Can Change the Number Even When Confounding Did Not
A practical guide to noncollapsibility of odds ratios for clinical researchers. Covers why crude and adjusted odds ratios can differ without confounding, when logistic regression invites over-interpretation, and what reviewers should demand instead.
Composite Endpoints: When One Trial Outcome Quietly Becomes Four Different Clinical Questions
A practical guide to composite endpoints for clinical researchers. Covers when endpoint bundles improve efficiency, when they distort clinical meaning, how soft components hijack results, and what reviewers should demand before trusting the headline.
Per-Protocol Effects: The Estimand Everyone Wants and the Bias Trap They Usually Build
A practical guide to per-protocol effects for clinical researchers. Covers sustained-adherence estimands, naive as-treated failure, selection bias after protocol deviation, and what reviewers should demand before trusting per-protocol claims.
Prediction vs Causation: Why Your Best Risk Model Still Cannot Tell You What to Treat
A practical guide for clinical researchers on the difference between prediction and causation. Covers why strong risk models do not identify treatment effects, how to frame the right estimand, and what reviewers should flag in AI-driven clinical studies.
Restricted Mean Survival Time: When Hazard Ratios Are Not the Clinical Answer
A practical guide to restricted mean survival time for clinical researchers. Covers what RMST estimates, when it beats the hazard ratio, how to choose the time horizon, and how to report results clinicians can actually interpret.
Landmark Analysis: Useful, Honest, and Frequently Overclaimed
A practical guide to landmark analysis for clinical researchers. Covers delayed treatment, immortal time bias, conditional survivor populations, landmark selection, and why a cleaner timeline still changes the causal question.
Targeted Maximum Likelihood Estimation: Doubly Robust, Not Doubly Forgiving
A practical guide to targeted maximum likelihood estimation for clinical researchers. Covers nuisance models, clever covariates, machine learning, overlap diagnostics, and why TMLE is robust in theory but never permission to stop thinking.
Case-Crossover Design: When Patients Become Their Own Controls
A practical guide to case-crossover designs for clinical researchers. Covers self-matching, hazard versus control windows, transient exposures, protopathic bias, time trends, and when this elegant design is exactly right or exactly wrong.
Multiple Imputation: Missing Data Does Not Become Innocent Because MICE Ran
A practical guide to multiple imputation for clinical researchers. Covers MICE, complete-case failure, MAR versus MNAR, imputation-model design, and why missing data needs causal thinking instead of software ritual.
G-Computation: Predict the Outcome Under Each Treatment Strategy
A practical guide to g-computation for clinical researchers. Covers counterfactual prediction, standardization, outcome modeling, positivity, model misspecification, and how to estimate causal effects by averaging predicted outcomes under competing interventions.
Estimands: The Causal Question You Should Define Before Running the Analysis
A practical guide to estimands for clinical researchers. Covers treatment strategies, intercurrent events, target populations, summary measures, and why many studies fail because they never define the actual causal question clearly.
Time-Varying Confounding: When Yesterday's Treatment Changes Today's Confounder
A practical guide to time-varying confounding for clinical researchers. Covers treatment-confounder feedback, why ordinary regression fails, and how MSMs, g-methods, and target trial logic handle evolving treatment decisions.
Informative Censoring: When Dropout Is Part of the Bias
A practical guide to informative censoring for clinical researchers. Covers loss to follow-up, treatment discontinuation, database exit, inverse probability of censoring weights, and why dropout can bias survival and causal estimates when it depends on prognosis.
Self-Controlled Case Series: When Each Patient Becomes Their Own Control
A practical guide to self-controlled case series for clinical researchers. Covers transient exposures, acute outcomes, fixed-confounding control, event-dependent exposure, and why within-person designs still live or die on timing assumptions.
Measurement Error: When Bad Variables Break Good Causal Methods
A practical guide to measurement error for clinical researchers. Covers noisy exposures, weak confounder proxies, surveillance-driven outcomes, validation strategies, and why sophisticated causal methods cannot rescue bad variables.
Competing Risks: When Kaplan-Meier Tells the Wrong Clinical Story
A practical guide to competing risks for clinical researchers. Covers death and discharge as competing events, why Kaplan-Meier can overstate event probability, and how cause-specific hazards and cumulative incidence answer different clinical questions.
Bias Amplification: When Adjustment Makes Unmeasured Confounding Worse
A practical guide to bias amplification for clinical researchers. Covers near-instruments, noisy severity proxies, treatment-prediction traps, and why the wrong adjustment variable can magnify residual confounding instead of reducing it.
Misclassification Bias: When Your Variables Lie Before the Model Starts
A practical guide to misclassification bias for clinical researchers. Covers wrong exposure and outcome labels, weakly measured confounders, surveillance-driven event detection, and why bad variables can distort causal estimates before modeling even begins.
Active Comparator New-User Design: The Observational Study Upgrade Most Drug Papers Need
A practical guide to the active comparator new-user design for clinical researchers. Covers why treated-versus-untreated comparisons fail, how new-user cohorts reduce prevalent-user bias, how active comparators narrow confounding by indication, and what reviewers should demand before trusting comparative effectiveness claims.
Selection Bias: When Your Study Sample Is the Problem
A practical guide to selection bias for clinical researchers. Covers referral filtering, survivor bias, complete-case analysis, informative loss to follow-up, collider-driven selection, and why a clean model cannot rescue a distorted sample.
Confounding by Indication: When Sicker Patients Make Treatments Look Dangerous
A practical guide to confounding by indication for clinical researchers. Covers treatment selection, severity-driven prescribing, contraindication bias, why routine adjustment often fails, and how to design observational comparisons that do not confuse prognosis with treatment effect.
Immortal Time Bias: The Fake Survival Advantage Hiding in Bad Study Design
A practical guide to immortal time bias for clinical researchers. Covers time zero, future-based exposure definitions, delayed treatment initiation, target trial emulation, and why you cannot adjust your way out of a broken timeline.
Interrupted Time Series: Strong Quasi-Experiments Need More Than a Before-and-After Plot
A practical guide to interrupted time series for clinical researchers. Covers level and slope changes, segmented regression, seasonality, autocorrelation, controlled ITS designs, and why a vertical line on a chart is not a causal estimate.
Parametric G-Formula: Estimating Causal Effects When Covariates Change Over Time
A practical guide to the parametric g-formula for clinical researchers. Covers time-varying confounding, dynamic treatment strategies, longitudinal simulation, model diagnostics, and why ordinary regression breaks when covariates are changed by prior treatment.
Proximal Causal Inference: What to Do When Unmeasured Confounding Is Still on the Table
A practical guide to proximal causal inference for clinical researchers. Covers proxy variables, treatment-inducing versus outcome-inducing proxies, bridge functions, completeness, and why this method is powerful but brutally assumption-heavy.
Positivity & Overlap: The Assumption Your Causal Estimate Cannot Survive Without
A practical guide to positivity and overlap for clinical researchers. Covers common support, extreme weights, trimming, overlap-focused estimands, and why many causal analyses fail because treated and untreated patients barely resemble each other.
Transportability & External Validity: When Your Causal Estimate Travels, and When It Absolutely Does Not
A practical guide to transportability and external validity for clinical researchers. Covers target populations, effect heterogeneity, trial selection, overlap, reweighting, and why “generalizable” is usually a lazy claim unless you prove the estimate can actually travel.
Collider Bias: How Adjustment Can Manufacture Associations
A practical guide to collider bias for clinical researchers. Covers common-effect conditioning, Berkson bias, selected cohorts, complete-case traps, and why the wrong adjustment set can literally manufacture a result.
Interference & Spillover Effects: When One Patient's Treatment Changes Another's Outcome
A practical guide to interference and spillover effects for clinical researchers. Covers SUTVA violations, direct versus indirect effects, partial interference, cluster and network designs, and why contamination is often the estimand trying to get your attention.
Principal Stratification: Estimating Effects When Post-Treatment Variables Matter
A practical guide to principal stratification for clinical researchers. Covers compliers, always-takers, truncation by death, latent strata, CACE/LATE, and why conditioning on observed post-treatment subgroups is usually causal self-sabotage.
Overadjustment Bias: When More Covariates Make Causal Inference Worse
A practical guide to overadjustment bias for clinical researchers. Covers mediators, colliders, post-treatment variables, propensity score misuse, and why the biggest adjustment set is often the least credible one.
Front-Door Criterion: The Causal Backdoor Alternative Nobody Uses Enough
A practical guide to the front-door criterion for causal inference. Covers full mediation, mediator-outcome confounding, identification logic, DAG requirements, and why most real datasets are nowhere near clean enough for a credible front-door design.
Negative Controls: The Bias Check Most Observational Studies Skip
A practical guide to negative control outcomes and exposures for clinical researchers. Covers residual confounding, selection bias, surveillance bias, falsification endpoints, and how to interpret a failed negative control without lying to yourself.
DAG Construction: How to Draw a Causal Graph Before You Touch the Model
A practical guide to DAG construction for clinical researchers. Covers time ordering, node selection, confounders vs mediators vs colliders, minimally sufficient adjustment sets, and why most covariate lists are just causal confusion wearing a regression badge.
Mediation Analysis: When You Want the Mechanism, Not Just the Effect
A practical guide to mediation analysis for clinical researchers. Covers direct and indirect effects, mediator-outcome confounding, treatment-induced confounding, interventional effects, and why most mediator-adjusted regressions are wrong.
G-Estimation: The Causal Method You Reach For When Time-Varying Confounding Breaks Regression
A practical guide to g-estimation and structural nested models for clinical researchers. Covers treatment-confounder feedback, blipped-down outcomes, identifying assumptions, and when g-estimation beats naive longitudinal regression or unstable weights.
Causal Forests: Finding Treatment Effect Heterogeneity Without Fooling Yourself
A practical guide to causal forests for estimating who benefits more, less, or not at all. Covers CATEs, honest splitting, overlap, validation, clinical use cases, and the reporting standards reviewers should expect.
E-values & Sensitivity Analysis: How to Stress-Test Causal Claims
A practical guide to the one question every observational study must answer: how strong would an unmeasured confounder have to be to erase your result? Covers E-values, confidence-limit interpretation, Rosenbaum bounds, negative controls, and the reporting language reviewers trust.
Mendelian Randomization: Using Genetics as Nature's Randomized Trial
How genetic variants serve as natural instruments for causal inference — and why horizontal pleiotropy, population stratification, and weak instruments break most published MR studies. Covers two-sample MR, MR-Egger, MR-PRESSO, and the STROBE-MR reporting checklist.
Marginal Structural Models: A Practical Guide for Clinical Researchers
How MSMs use stabilized inverse probability weights to handle time-varying confounders — the ones that change over time and are affected by prior treatment. Covers weight estimation, model fitting, clinical examples, and common pitfalls.
Structural Causal Models & DAGs: A Practical Guide for Clinical Researchers
The causal framework behind every method you use. Covers DAGs, d-separation, do-calculus, backdoor/frontdoor criteria, mediation analysis, and how to draw the graph that makes your analysis work.
Inverse Probability Weighting: When PSM Discards Your Data
Why IPW outperforms matching by keeping all patients — and how extreme weights, positivity violations, and wrong variance estimators break published analyses silently.
Double Machine Learning: A Practical Guide for Clinical Researchers
How DML uses machine learning to estimate causal effects while controlling for high-dimensional confounders. Covers cross-fitting, Neyman orthogonality, clinical applications, and implementation in EconML.
Target Trial Emulation: A Practical Guide for Clinical Researchers
The framework that bridges observational data and causal claims — by asking what RCT you wish you had. Covers protocol specification, time zero alignment, clone-censor-weight, immortal time bias, and reporting.
Regression Discontinuity Design: A Practical Guide for Clinical Researchers
RDD turns arbitrary thresholds into causal evidence. Covers sharp vs fuzzy designs, bandwidth selection, manipulation testing, clinical applications, and a complete reporting checklist.
Synthetic Control Methods: Building Counterfactuals When DID Fails
How to construct a synthetic twin from donor pools when parallel trends don't hold. Covers SCM optimization, validation via placebo tests, modern extensions (ASCM, SDID), and common pitfalls.
Difference-in-Differences: A Practical Guide for Clinical Researchers
When and how to use DID in clinical research. Covers parallel trends, staggered adoption, common pitfalls, reporting checklist, and modern estimators.
Instrumental Variables: When Observational Data Meets Unmeasured Confounding
When PSM and regression fail because of unmeasured confounding, IV methods offer a way forward. A practical guide covering instruments, LATE, Mendelian randomization, and the exclusion restriction.
Propensity Score Matching: A Practical Guide for Clinical Researchers
What PSM actually does, when it fails, and how to report it correctly. Written for researchers who want to use it — not just cite it.