Skip to content
Subconscious

UserTesting, Maze, Lookback, or Causal AI: Which Method Fits the Decision?

Use usability research to observe people interacting with a product or prototype. Use a controlled comparison when selecting among designs, prices, messages, or concepts requires an intervention estimate. These capabilities can overlap within UserTesting, Maze, or other platforms; inspect the configured task and endpoint.

These methods are complements. Each becomes expensive when asked to answer a question it cannot measure.

Discover the problem, choose an observed task or choice endpoint, run an appropriate controlled comparison, and evaluate deployment.
Usability platforms can support controlled comparisons; the measured endpoint determines the claim.

What does each method measure?

MethodWhat it measuresBest questionMain limitation
UserTesting and UserZoom workflowsParticipant behavior and commentary; supported studies can compare randomized task variantsWhere does the experience fail, or how do variants affect a measured task?Prototype or elicited results do not automatically establish market performance
MazeTask completion, prototype interaction, survey, and usability outcomesCan people complete a defined task?Task success is not the same as market choice
LookbackLive or recorded interface behavior plus participant explanationWhy did the participant act that way in the interface?Small qualitative sessions do not estimate a causal market effect
SubconsciousChoice responses to controlled commercial alternatives in simulated audiences, recruited human participants, or bothWhich defined action produces the stronger behavioral response?Does not replace session-level observation of interface use

When should you use UserTesting?

UserTesting records real people interacting with a product, letting researchers watch sessions, hear participants explain their actions, and see where an interface creates confusion.

UserTesting's UserZoom task documentation includes assigning participants to website or prototype versions. Randomization availability depends on study type. Define whether the endpoint is preference, task completion, or actual deployed behavior before interpreting the contrast.

Request current quotes for participant eligibility, recruitment, study volume, seats, supported design, analysis, and services. Obtain a delivery estimate for the actual audience and instrument; the unavailable historical annual-contract examples do not establish current price or timing.

When should you use Maze or Lookback?

Maze's current variant-comparison documentation describes exclusive random assignment and randomized presentation order. It distinguishes research comparison from live production A/B testing. Check plan availability and specify the metric before the study.

Lookback supports live and recorded usability sessions. It fits a researcher who needs follow-up questions, screen interaction, participant commentary, and direct observation.

Both retain an advantage causal simulation should not pretend to replace: a researcher can watch a real participant use the interface.

When Subconscious is the right choice

The study starts with an intervention and an outcome. The team defines the audience, holds the decision context stable, and compares plausible actions.

A synthetic choice comparison estimates a contrast in generated responses. For a human check, align population, alternatives, outcome, and analysis where feasible, and record task differences. Actual interface completion needs participants using the interface.

This is the right method when a product or marketing leader must choose:

The defined choice comparison can inform action selection within the study conditions. It does not by itself establish usability completion, purchase behavior, or launch performance.

The supported causal experiment use cases show where this method fits. The research program explains why human baselines and validation remain part of the workflow.

Use the methods in sequence

  1. Use exploratory interviews to find the problem and learn the language customers use.
  2. Use an observed interface study to diagnose friction or compare task variants under an appropriate design.
  3. Compare commercial alternatives on a defined choice endpoint, with relevant human validation for a synthetic study.
  4. Observe the rollout to learn whether the measured effect holds in market.

It prevents a stated preference, a usability failure, and a causal effect from being treated as the same result.

Limitations

Remote usability tools depend on recruitment quality, task design, and the realism of the tested interface. A Subconscious experiment, whether simulated or human, depends on the study design, audience definition, alternatives, and outcome chosen for that decision.

No method removes the need for human judgment. Consequential product, health, financial, or policy decisions require real-world validation and qualified review.

If the team has a concrete action to compare rather than an interface task to observe, bring the decision to a Subconscious working session.