A decision-relevant assumption has been selected and its origin is traceable.
Test Card
A practical method guide: clarify prerequisites and materials, understand the complete process, and look up outputs, pitfalls, and stop criteria.
Prerequisite
What needs to be finished first
The audience, test environment, data capture, and accountable person are broadly available.
Preparation
What needs to be ready before start
Test Card, prioritized assumption, available baseline, audience access, and information about privacy, ethics, and operational risks.
Test owner · Product · Research or Analytics · privacy or risk expert where needed
Prepare the prioritized assumption, upcoming decision, available audience, known baseline, and relevant safeguard boundaries.
20 to 35 minutes for planning; execution is separate
Arrange the four fields visibly. Work through hypothesis, test, metric, and success threshold in that order, then freeze the card before launch.
Core question
The one question this method answers
How will we test this critical assumption with a clear test, a suitable metric, and a success threshold set in advance?
Flow
Phases with goal, actions, output and pitfall
The Test Card is a compact planning artefact for one test. It connects hypothesis, test, metric, and success threshold before execution. This lets the team later determine whether the evidence sufficiently supports the assumption or calls for another learning step.
Test Card
Fictional worked example
Four fields connect one critical assumption to a test, metric, and success threshold before execution.
Hypothesis
What must be true for the idea?
Project leads activate a weekly risk report when setup takes no more than two minutes.
Test
How will we test the hypothesis?
40 active project leads see a clickable setup dialogue for seven days.
Metric
Which observation counts as evidence?
Share of exposed project leads who complete the setup.
Success threshold
Which boundary applies before launch?
At least 12 completions from 40 valid exposures; stop after one privacy complaint.
- 1
Hypothesis
State the critical assumption

Goal
Define one specific, falsifiable claim that matters to the next decision.
Inputs for this step
- Prioritized assumption
- Product or business model context
- Upcoming decision
Actions
- State what must be true for the idea to work.
- Name the audience, expected behaviour, and relevant context.
- Limit the statement to exactly one testable assumption.
Output
Testable hypothesis
Audience · expected behaviour · context
Quality check
Can an observable result clearly weaken the claim?
Pitfall
The statement bundles several assumptions or merely describes a desired solution. Reduce the statement to one decision-relevant assumption and add observable behaviour.
Timing: 5 minutes
- 2
Test
Choose the smallest meaningful test

Goal
Design a responsible procedure that produces evidence specifically for the hypothesis.
Inputs for this step
- Testable hypothesis
- Access to the audience or test environment
- Time, cost, privacy, and risk constraints
Actions
- Choose the smallest intervention with sufficient evidence strength.
- Specify participants, procedure, duration, and data capture.
- Check consent, privacy, deception risks, and stop authority.
Output
Executable test
Participants · intervention · procedure · duration · safeguards
Quality check
Does the test produce evidence that directly fits the hypothesis?
Pitfall
The test measures convenient opinions although the hypothesis concerns actual behaviour. Match the test signal to the hypothesis and observe behaviour where required.
Timing: 10 minutes
- 3
Metric
Define the observable signal

Goal
Specify which observation or calculation counts as evidence and how it will be captured.
Inputs for this step
- Hypothesis and test procedure
- Available data source
- Relevant observation window
Actions
- Choose one primary metric that represents the expected behaviour.
- Specify numerator, denominator, target segment, and window where the metric is calculated.
- Set the data source, capture point, and quality check.
Output
Unambiguous metric
Signal · data source · calculation · window
Quality check
Can two people derive the same value from the same raw data?
Pitfall
The metric counts activity without representing the behaviour claimed in the hypothesis. Align the metric with the claimed behaviour and record its calculation and data source.
Timing: 5 to 10 minutes
- 4
Success threshold
Set the threshold in advance

Goal
Define a quantitative or clearly observable boundary at which the result is considered sufficient.
Inputs for this step
- Defined metric
- Baseline or justified reference value
- Consequence of the later decision
Actions
- Set the success threshold before viewing results.
- Add minimum data or required observation quality.
- Record safety, ethics, or data-quality boundaries for pausing and stopping.
Output
Pre-set success threshold
Boundary · minimum data · pause and stop limits
Quality check
Can the team determine whether the threshold was met without introducing a new interpretation?
Pitfall
The threshold remains vague or is selected to fit the observed result. Justify, date, and freeze the boundary and minimum data before launch.
Timing: 5 to 10 minutes
Artifact
What comes out at the end
A versioned Test Card as a compact test plan. Record results and conclusions in a separate Learning Card after execution.
Retain date, owner, hypothesis source, and status for each card. Record changes after launch as a new version.
- Shared paper or digital Test Card
- Data or observation protocol for later execution
- Decision log for approval and accountability
test-card-working-template.md
Plan one decision-relevant test in a compact session. Execution and evaluation receive separate dates and artefacts.
Test Card: Worksheet
Question
To be confirmed
Desired outcome
To be confirmed
Scope
Complete Test Card Time: 20 to 35 minutes
Complete this worksheet on paper or in your own document during the work.
Preparation for this scope
Method setup
- Materials: Test Card, prioritized assumption, available baseline, audience access, and information about privacy, ethics, and operational risks.
- Roles: Test owner · Product · Research or Analytics · privacy or risk expert where needed
- Advance information: Prepare the prioritized assumption, upcoming decision, available audience, known baseline, and relevant safeguard boundaries.
- Overall time needed: 20 to 35 minutes for planning; execution is separate
- Setup: Arrange the four fields visibly. Work through hypothesis, test, metric, and success threshold in that order, then freeze the card before launch.
Prepare for the work steps
Hypothesis: State the critical assumption
- Prioritized assumption
- Product or business model context
- Upcoming decision
Test: Choose the smallest meaningful test
- Testable hypothesis
- Access to the audience or test environment
- Time, cost, privacy, and risk constraints
Metric: Define the observable signal
- Hypothesis and test procedure
- Available data source
- Relevant observation window
Success threshold: Set the threshold in advance
- Defined metric
- Baseline or justified reference value
- Consequence of the later decision
Test Card
Hypothesis
What must be true for the idea?
...
Test
How will we test the hypothesis?
...
Metric
Which observation counts as evidence?
...
Success threshold
Which boundary applies before launch?
...
Strategyzer: Validate Your Ideas with the Test Card
Hypothesis: State the critical assumption
Expected artifact: Testable hypothesis
Audience · expected behaviour · context
Entry:
...
- Can an observable result clearly weaken the claim?
Test: Choose the smallest meaningful test
Expected artifact: Executable test
Participants · intervention · procedure · duration · safeguards
Entry:
...
- Does the test produce evidence that directly fits the hypothesis?
Metric: Define the observable signal
Expected artifact: Unambiguous metric
Signal · data source · calculation · window
Entry:
...
- Can two people derive the same value from the same raw data?
Success threshold: Set the threshold in advance
Expected artifact: Pre-set success threshold
Boundary · minimum data · pause and stop limits
Entry:
...
- Can the team determine whether the threshold was met without introducing a new interpretation?
Open questions and next steps
...
Method guide: https://methodatlas.meierhoff-systems.de/en/methods/test-card/run-sheet
Example output
Concrete filled scenario, fictional example
Fictional example for orientation. Counts and timings describe this case and are not universal requirements.
test-card-beispiel.md
Concrete filled scenario, fictional example
Fictional example: we believe project leads will activate a weekly risk report when setup takes no more than two minutes. Test: 40 active project leads see a clickable setup dialogue for seven days. Metric: share of exposed project leads who complete setup. Success threshold: at least 12 completions from 40 valid exposures. Pause if event capture fails; stop after one privacy complaint.
Pitfalls
Recognize symptoms and steer against them
Quick overview, the full text is in the matching phase in the flow
- State the critical assumptionThe statement bundles several assumptions or merely describes a desired solution.View in phase
- Choose the smallest meaningful testThe test measures convenient opinions although the hypothesis concerns actual behaviour.View in phase
- Define the observable signalThe metric counts activity without representing the behaviour claimed in the hypothesis.View in phase
- Set the threshold in advanceThe threshold remains vague or is selected to fit the observed result.View in phase
State the critical assumption
The statement bundles several assumptions or merely describes a desired solution.
Reduce the statement to one decision-relevant assumption and add observable behaviour.
Choose the smallest meaningful test
The test measures convenient opinions although the hypothesis concerns actual behaviour.
Match the test signal to the hypothesis and observe behaviour where required.
Define the observable signal
The metric counts activity without representing the behaviour claimed in the hypothesis.
Align the metric with the claimed behaviour and record its calculation and data source.
Set the threshold in advance
The threshold remains vague or is selected to fit the observed result.
Justify, date, and freeze the boundary and minimum data before launch.
Stop criteria
When to pause, resolve prerequisites, or choose another approach
Finished the method guide?
Go to the profile for purpose, similar methods, and sources or continue to the next method in the catalog.

