Manuel B. Garcia

Manuel B. Garcia serves as the Senior Director for Educational Technology and Digital Learning at FEU Institute of Technology, Manila, Philippines. Read More

Contact Info

1607, FEU Tech Building,
P. Paredes St, Sampaloc,
Manila, Philippines
mbgarcia@feutech.edu.ph

Follow Me

How Do You Synthesize Evidence Across Age Groups?

Evidence from different age groups can often contribute to the same synthesis, but age differences should not be ignored when they could alter baseline risk, intervention effects, measurement, or the phenomenon itself.

569
Synthesizing Evidence Across Age Groups Guide 569 of 899
01 · The Question

Can evidence from children, younger adults, and older adults really be combined?

Your review question concerns a broad population, but the available studies do not divide neatly by age. One study includes adolescents. Another includes adults aged 18 to 65. A third focuses on people over 60. Several report only the mean age of participants.

Should these studies contribute to one synthesis, or does combining them risk hiding age-related differences?

Age can matter substantially, but an age difference does not automatically make evidence incomparable. The important question is whether age could plausibly change the outcome, effect, mechanism, measurement, or applicability relevant to your review.

02 · The Short Answer

Combine age groups when they answer the same substantive question

In Brief

Evidence across age groups can be synthesized when the studies address a sufficiently common question and there is no compelling reason to expect age differences to make their findings substantively incomparable.

If age could plausibly modify the finding, preserve age information and investigate it where the evidence allows. Do not assume that different ages require separate syntheses, but do not assume that an overall average applies equally to every age group either.

03 · What You Need to Know

Age can matter in several different ways

Age is often treated as a routine demographic variable, yet its methodological meaning depends on the research question. It may represent biological development, accumulated exposure, disease risk, social roles, educational stage, cognitive development, treatment tolerance, technology use, or numerous other characteristics.

That makes “Does age matter?” too broad a question. You need to ask how age could matter for the specific phenomenon being synthesized.

Age difference does not automatically create indirect evidence

GRADE guidance treats population differences, including age differences, as a potential source of indirectness when researchers apply evidence to a target population. Yet differences between the studied and target populations do not automatically justify lower confidence. The concern becomes important when there is a credible reason to expect those differences to change the effect relevant to the target question.

For example, evidence from adults may sometimes inform decisions about older adults if the underlying mechanism and relative effect are expected to remain similar. In another context, evidence from adults may provide poor guidance for young children because developmental physiology, dosing, measurement, behavior, or implementation differs fundamentally.

The substantive question determines which situation you face.

Distinguish baseline risk from effect modification

An age group can have a different baseline probability of an outcome without necessarily experiencing a different relative intervention effect.

Suppose an adverse event is much more common among older adults. Even if an intervention produces the same relative risk reduction across ages, the absolute benefit may be larger among older adults because their starting risk is higher.

Prognostic effect Age is associated with the likelihood of the outcome regardless of the intervention or exposure.
Effect modification The effect of the intervention or exposure itself differs according to age.

Confusing these concepts can lead researchers to conduct age subgroup analyses simply because age predicts the outcome. A prognostic factor is not necessarily an effect modifier.

Do not create age categories mechanically

Age is fundamentally continuous, even though studies frequently convert it into categories such as children, adolescents, younger adults, middle-aged adults, and older adults. Those categories may be substantively useful, but their boundaries should not be treated as natural laws.

A threshold of 60 or 65 years may make sense for one research question and little sense for another. Developmental research may require much narrower categories, while another phenomenon may change gradually across adulthood.

When possible, base age categories on biological, clinical, developmental, social, or policy relevance rather than selecting cut points merely because they make subgroup analysis convenient.

Broad age ranges can hide within-study variation

A study reporting a mean participant age of 45 years might include people aged 18 to 80. Another study with the same mean could contain almost exclusively participants aged 40 to 50. Their reported means therefore do not establish equivalent age composition.

This becomes particularly important in meta-regression. Cochrane guidance highlights age as an example of aggregation bias: a true relationship between age and treatment effect may exist within individual studies but remain invisible when researchers compare only the mean ages of entire studies.

The reverse problem can also occur. A relationship between study-level mean age and effect estimates does not prove that age modifies the effect among individuals.

Watch Out

Do not interpret a study-level meta-regression of mean age as if it were an individual-participant analysis. Study averages discard within-study information and can produce ecological or aggregation bias.

Ask whether the intervention operates differently by age

Age deserves greater attention when there is a plausible reason that the intervention or exposure would operate differently across the age range.

In clinical research, pharmacokinetics, comorbidities, organ function, developmental physiology, treatment tolerance, or competing risks may matter. In education, developmental stage, prior knowledge, autonomy, and curriculum can change the meaning of an intervention. In technology research, device familiarity or patterns of use may differ across generations.

None of these differences should simply be presumed. They provide reasons to investigate whether age-related effect modification is plausible.

Check whether the outcome has the same meaning across ages

Measurement comparability can become as important as intervention comparability. The same instrument may not have equivalent validity across children and adults. A behavioral outcome may have different developmental meanings. Functional independence, academic performance, social participation, or quality of life may also be operationalized differently across life stages.

If outcomes are not measuring sufficiently comparable constructs, pooling them because they share a label can produce a deceptively precise answer to an unclear question.

Compare effects directly rather than comparing significance

A familiar subgroup mistake occurs when researchers observe a statistically significant effect among younger participants and a nonsignificant effect among older participants and conclude that the treatment works only in younger people.

That conclusion does not follow. Statistical significance depends partly on sample size and precision. The effect estimates themselves might be very similar.

To evaluate age-related effect modification, the relevant analysis examines evidence for a difference between effects. Even then, subgroup findings should be interpreted alongside their precision, prespecification, biological or substantive plausibility, consistency across studies, and vulnerability to confounding or multiple testing.

Sometimes individual participant data provide a better answer

When age-related modification is central to the research question, aggregate study-level data may be inadequate. Individual participant data meta-analysis can allow age to be modeled at the participant level rather than relying on study means or broad published categories.

This can be particularly useful when age is continuous or when different studies use incompatible age cutoffs. It does not remove every source of bias, but it can avoid some limitations inherent in study-level age comparisons.

Age also affects applicability

Even when a synthesis is statistically coherent, you still need to ask whether it applies to the population of interest. A body of evidence dominated by middle-aged adults may offer less direct evidence for children or very old adults when age-related differences could plausibly alter the finding.

This is part of the broader question of when evidence from another population is directly relevant to yours. The answer depends on the characteristics that matter to the phenomenon, not simply whether the populations have different demographic labels.

04 · A Practical Example

Suppose an intervention is studied from adolescence through older adulthood

Hypothetical Example

A digital behavior-change intervention across age groups

Imagine a hypothetical review of a digital program intended to increase physical activity. Fifteen studies include participants ranging from adolescents to adults over 70.

Define the common question All studies compare the digital intervention with a broadly comparable control condition and measure change in physical activity.
Identify why age could matter The reviewers prespecify age as potentially relevant because device use, mobility limitations, baseline activity, and intervention engagement may vary across the age range.
Inspect the data Several studies include broad overlapping age ranges. Only a minority report effects separately for predefined age groups.
Avoid a weak shortcut The reviewers do not classify entire studies as “young” or “old” solely from mean participant age and then interpret the resulting difference as individual-level effect modification.
Synthesize cautiously They report the overall evidence while preserving available age-specific estimates. Any apparent age pattern is described according to the strength and limitations of the evidence rather than converted into a definitive age threshold.

The result may be less dramatic than announcing that an intervention “works only below age 60,” but it is methodologically more defensible. Age thresholds can look wonderfully crisp in a forest plot while the underlying biology remains stubbornly continuous.

05 · What Researchers Often Get Wrong

Common mistakes when synthesizing evidence by age

Misconception

Different age groups must always be synthesized separately

Age differences matter only to the extent that they change the substantive question, effect, outcome, or applicability. Separating every study by age can fragment evidence unnecessarily when no credible age-related modification is expected.

Misconception

If older people have worse outcomes, the intervention must work differently for them

Higher baseline risk does not establish effect modification. Older participants may experience different absolute outcomes even when the relative effect of an intervention remains similar across ages.

Misconception

A significant effect in younger adults but not older adults proves an age interaction

Separate significance tests do not test whether effects differ. Compare the effects directly using an appropriate interaction or subgroup-difference analysis and interpret the result with its uncertainty.

Misconception

The mean age tells you which age group a study represents

A mean conceals the distribution. Studies with identical mean ages can contain very different age ranges, and assigning an entire study to an age category from its mean can create misleading study-level comparisons.

Misconception

A statistically significant age meta-regression proves age modifies the effect

Study-level meta-regression is observational and vulnerable to confounding, aggregation bias, and ecological inference. It can support a hypothesis about age-related heterogeneity, but it does not by itself establish an individual-level age effect.

06 · What This Means for You

Let the research question determine how much age should structure the synthesis

Before dividing evidence into age groups, specify why age could change the finding. That reasoning should guide extraction, subgroup definitions, synthesis, and interpretation.

A simple decision framework

If age differs across studies but no plausible age-related effect modification is expected
A combined synthesis may remain appropriate, provided the studies are otherwise sufficiently comparable.
If developmental, biological, behavioral, or contextual mechanisms make age potentially important
Preserve age-specific information and prespecify appropriate analyses where possible.
If studies report usable age-specific effect estimates
Compare the estimates directly rather than relying on separate within-group significance tests.
If only study-level mean ages are available
Treat meta-regression cautiously and avoid translating study-level associations directly into individual-level conclusions.
If the evidence scarcely represents your target age group
Consider whether the evidence is indirect rather than silently assuming that effects generalize.

Sometimes the most useful conclusion is that the available evidence does not permit a reliable age-specific inference. That is not a failed synthesis. It identifies precisely where the evidence stops supporting confident interpretation.

If the target age group differs substantially from those studied and there is a credible reason the finding could change, it may be necessary to state explicitly that the evidence does not generalize directly.

07 · A Quick Checklist

Before combining evidence across age groups, check:

Before synthesizing across ages, check:
Does age have a plausible biological, developmental, behavioral, or contextual relationship with the finding?
Are you distinguishing age-related baseline risk from genuine effect modification?
Do the intervention, exposure, comparator, and outcome have comparable meanings across the age range?
Are age categories substantively justified rather than selected merely for analytical convenience?
Have you examined age ranges and distributions rather than relying only on study means?
Are subgroup conclusions based on direct comparisons between effects rather than differences in statistical significance?
Could study-level age analyses suffer from confounding, ecological bias, or aggregation bias?
Is your target age group adequately represented in the evidence?
08 · Frequently Asked Questions

Questions about synthesizing evidence across ages

Can studies of children and adults be combined in one meta-analysis?

Sometimes. The decision depends on whether the intervention or exposure, outcome, underlying mechanism, and research question remain sufficiently comparable. Developmental differences may make separate synthesis necessary for some questions but not others.

Should I automatically separate adults aged 65 and older?

No. A cutoff such as 65 may be useful in some contexts, but it is not a universal biological boundary. Choose age categories according to the phenomenon, population, established disciplinary conventions, and your prespecified research question.

Can I use mean age in a meta-regression?

Yes, but interpretation requires caution. Mean age is a study-level characteristic and may conceal substantial within-study variation. An association between mean age and effect size across studies does not establish individual-level effect modification by age.

What if studies use different age categories?

Avoid pretending incompatible categories are identical. Where possible, harmonize categories only when their underlying populations can be mapped defensibly. If age-specific inference is central and aggregate reports are inadequate, individual participant data may provide a better basis for analysis.

Does a larger absolute benefit in older adults mean treatment works better for them?

Not necessarily. Higher baseline risk can produce a larger absolute benefit even when the relative treatment effect is similar across age groups. Specify which effect measure you are comparing before interpreting age differences.

What if almost all studies include middle-aged adults?

The overall synthesis may still answer a question about that population, but applying it to children, adolescents, or very old adults may involve indirectness when age-related differences could plausibly alter the effect or outcome.

09 · The Bottom Line

Age should structure the synthesis only when it matters to the question

The Bottom Line

Synthesize evidence across age groups when the studies address a sufficiently common question, but preserve age-specific evidence when there is a credible reason that age could modify the effect, outcome, mechanism, or applicability.

Do not confuse baseline risk with effect modification, arbitrary age categories with natural boundaries, or study-level mean age with individual-level evidence. When age genuinely matters, make it visible. When the evidence cannot establish an age difference, say that rather than manufacturing one from subgroup statistics.

10 · Sources and Further Reading

Sources and further reading

11 · Cite this Guide

How to Cite This Guide

This guide is intended to be read, shared, and used in research, teaching, and academic work. If you draw on its ideas, explanations, or other content, please acknowledge the source by citing the guide. Doing so gives appropriate credit and helps your readers locate the original resource.

Has the Field Guide helped your research?

If a guide helped clarify a question, inform a research decision, or move your work forward, I would love to hear about your experience. Your story may also help other researchers discover the Field Guide.

Share Your Experience
Takes only a few minutes