Site Logo
Research MethodsEvidence Synthesis

Meta-synthesis: integrating qualitative findings across studies

Evidano6 min read

Meta-synthesis does for qualitative research what meta-analysis does for trials — with the crucial difference that what gets combined is interpretations, not numbers. The synthesist gathers the qualitative studies on a question, treats their findings as data, and builds an interpretation that goes beyond any single study: not a summary of summaries but a new, higher-order account. The method’s recurring tension is depth versus aggregation — every included study flattened a lived reality into themes once already, and synthesis flattens again. Good meta-synthesis manages that loss deliberately; bad meta-synthesis produces theme soup.

First-, second-, and third-order constructs

The method’s working vocabulary: first-order constructs are participants’ own accounts (the quotes in primary papers); second-order constructs are the primary authors’ interpretations (their themes); third-order constructs are the synthesist’s interpretation of those interpretations.

The synthesist mostly works with second-order material — which is why the quality and reporting depth of the primary studies bounds everything. A primary paper that presents three thin themes and two quotes contributes almost nothing synthesisable, however large its sample.

Keeping the three levels distinct in the analysis is what separates synthesis from paraphrase: a third-order construct must be traceable down through named second-order themes to first-order evidence.

The methodological anchors

Thematic synthesis — the most used approach — was formalised by Thomas and Harden in Methods for the thematic synthesis of qualitative research in systematic reviews: line-by-line coding of findings, descriptive themes, then analytic themes that go beyond the primary studies.

The broader craft — sampling, appraisal, and the varieties of synthesis — is laid out in Sandelowski and Barroso’s Handbook for Synthesizing Qualitative Research (Springer Publishing), still the fullest treatment of the judgement calls involved.

Confidence in synthesised findings now has a formal system: GRADE-CERQual, introduced by Lewin and colleagues in Using qualitative evidence in decision making for health and social interventions, which grades each synthesised finding on methodological limitations, coherence, adequacy, and relevance.

How a thematic meta-synthesis proceeds

Frame the question and the sampling logic

Decide whether the review is exhaustive (all qualifying studies) or purposive (studies sampled for conceptual contribution, common when hundreds qualify). Both are legitimate; the choice must be declared and its rationale defended.

Search and appraise for thickness, not just conduct

Systematic searching applies, with the known caveat that qualitative work is indexed inconsistently. Appraisal (CASP is typical) serves two roles: flagging weak conduct and — more decisive in practice — identifying thin reporting, since a study’s findings can only contribute what its paper actually presents.

Code the findings line by line

Everything labelled results or findings in each primary paper — themes, explanations, and the quotes they carry — is coded freely, staying close to the primary authors’ meanings. Codes accumulate and consolidate across studies.

Build descriptive, then analytic themes

Descriptive themes organise the coded material while remaining faithful to the primaries. The analytic step then asks the synthesist’s own question of that structure — what explains the pattern? what do these accounts, together, show that none shows alone? This is where third-order constructs are made, and where the method either earns its keep or stops at summary.

Grade and present each synthesised finding

State each finding with the studies behind it, its CERQual confidence grade, and the reasons for that grade. A findings table with per-finding provenance is now the expected format in health and policy contexts.

Worked example: adherence to tuberculosis treatment

A synthesis of qualitative studies on TB treatment adherence included 28 papers from 14 countries. Line-by-line coding of the findings sections produced 214 codes, consolidated into eleven descriptive themes — stigma, cost, side effects, provider relationships, and so on — all recognisable from the primary studies.

The analytic step produced the third-order construct that made the synthesis cited: adherence behaviour tracked the visibility of treatment, not its burden. Clinic queues visible to neighbours, workplace absences, packaging identifiable at home — across settings, patients managed exposure first and regimen second. No single primary study had stated this; seven had second-order themes consistent with it, and the deviant cases (two studies where treatment was home-delivered and discreet, with high adherence despite severe side effects) strengthened rather than broke the construct.

Each synthesised finding was CERQual-graded; the visibility construct carried high confidence (coherent across 19 studies of adequate thickness), while a finding about gendered adherence differences was graded low — four studies, two with serious limitations — and flagged as a research gap rather than a conclusion.

Common mistakes

  • Stopping at descriptive themes. A list of themes that any included study already contained is aggregation, not synthesis.
  • Untraceable third-order claims. Every synthesised construct needs its chain of custody down to primary evidence.
  • Treating study count as confidence. Twelve thin studies can ground less than four thick ones; CERQual exists precisely to replace counting.
  • Synthesising across incompatible questions. Studies about experiencing a disease and studies about experiencing its treatment are adjacent, not identical; forcing them together produces themes true of neither.
  • Ignoring reflexivity. The synthesist’s framing shapes the third-order constructs; the write-up should say who read what and how disagreements were handled.

Limitations

Meta-synthesis inherits and compounds the flattening of primary analysis: participants’ words arrive twice interpreted, and context — the soul of qualitative work — survives only as much as primary reporting allowed. The method reads best as building transferable mid-range theory, not as recovering experience.

It is also hostage to reporting conventions: word limits in primary journals systematically thin the material available to synthesise, and the synthesist cannot repair what was never printed.

And the interpretive step resists standardisation by design. Two teams can code identically and build different third-order constructs; transparency about the chain of interpretation is the method’s honesty mechanism, not a guarantee of convergence.

Where software helps

Line-by-line coding of 28 findings sections is exactly the corpus-scale coding work AI does well: Evidano codes the findings passages inductively or against your developing frame and links every code to its source passage, so descriptive themes are built on retrievable evidence and the traceability chain CERQual asks for exists by construction.

The analytic themes — the third-order move — are the synthesist’s interpretation and should stay that way; the tool’s role is to make the second-order layer completely searchable while you make the leap.

Topics

  • meta-synthesis
  • qualitative meta-synthesis
  • qualitative evidence synthesis
  • thematic synthesis
  • CERQual
  • third-order constructs

Other methods in evidence synthesis

Written guides are linked directly; the rest have a reference entry in the methodology directory.

Published research using these methods

Studies and evaluations where this family of method was applied with Evidano — the work, not the claim.

Keep reading

Browse all articles