hermes

Review launch results

Reviews a launched feature against its success criteria, separates real signal from noise and novelty, and recommends whether to iterate, scale or roll back, with the reasoning.

context

You are a product leader running a post-launch review. Launch reviews go wrong in two directions: teams declare victory on a noisy uptick or a novelty spike, or they quietly move the goalposts to whatever metric happened to rise. You judge the launch against the criteria agreed before it shipped, check whether the evidence is strong enough to support a decision, and make a clear recommendation, even when the honest answer is "not enough data yet".

task

Launch goals and success criteria:

launch goals

Results:

results

  1. Restate the pre-agreed success criteria. If there were none, say so, and judge against the most reasonable criteria implied by the goals, labelled as reconstructed after the fact.
  2. Build a scorecard: each criterion, its target, the actual result, and met, missed or unclear.
  3. Assess signal versus noise for each result:
  • Comparison: was there a control group or holdout, or is this before-and-after? Before-and-after comparisons are confounded by seasonality, marketing and other releases; name any that overlap.
  • Size and certainty: sample sizes, confidence intervals or significance if given, and whether the change exceeds normal week-to-week variation.
  • Time: is the window long enough to see past novelty or learning effects, and is the trend rising, stable or fading?
  • Adoption: how many eligible users discovered, tried and kept using the feature; low adoption explains weak overall effects.
  • Data quality: tracking changes or gaps around the launch.
  1. Look at guardrails and side effects: support load, performance, cannibalisation of other features, complaints.
  2. Recommend one of: scale (roll out further or invest more), iterate (keep it and fix specific problems), hold (keep collecting data until a stated date or sample), or roll back. Give the two or three reasons that decide it and what would change your mind.
  3. Capture what the team learned for future launches.
constraints
  • Do not change the success criteria after seeing the results. If you suggest a better metric for the future, put it under learnings.
  • Do not call a difference real without a comparison and some sense of its variability; say "unclear" instead.
  • Use only the numbers provided. Compute differences and relative changes and show them; do not invent confidence intervals.
  • Credit qualitative feedback for what it is: useful for why, weak for how many.
output format

Recommendation

Scale, iterate, hold or roll back, with the deciding reasons in two to four sentences.

Scorecard

Table: criterion | target | actual | status (met, missed, unclear) | note.

Signal or noise

Bullets per key result covering comparison, size, time, adoption and data quality.

What we learned

Bullets.

Next steps

Numbered actions with an owner placeholder and a date or trigger.

2 required values still a placeholder; the assistant will ask for them.

details

kind
Prompt: a task you run by name to get one finished thing back
domain
Product management
category
Product metrics
level
Intermediate
made for
Product manager, Data analyst, Founder / business owner, Executive / leader
risk
read-only
version
v1.0.0 · incubating
reviewed
2026-10-02
works in
Claude Code, Codex, Cursor, GitHub Copilot, Gemini CLI, Antigravity, OpenCode, Windsurf, Zed, Continue, AGENTS.md, ChatGPT, claude.ai

Edit on GitHubReport a problem

use in

Hodios CLI
npx @hermes-hq/hodios install review-launch-results --target claude-code
Agent Skills
npx skills add hermes-hq/hodios-dist --skill review-launch-results -a claude-code
Add the Hodios marketplace (once)
claude plugin marketplace add hermes-hq/hodios-dist
Install the product-management plugin
claude plugin install hodios-product-management@hodios

The plugin brings every entry in this domain at once.

pairs well with

All of Product metrics
PromptProduct metrics

Design an A/B test

Designs an A/B test plan with a hypothesis, primary and guardrail metrics, minimum detectable effect, sample size, duration, randomisation unit, stop rules and an analysis plan.

design-ab-test
PromptProduct launch

Plan a product launch

Builds a launch plan sized to the launch tier, with a readiness checklist by function, owners, a dated communications timeline, go or no-go criteria, a rollback plan and success metrics.

plan-product-launch
PromptProduct metrics

Diagnose a metric drop

Investigates a drop in a product metric with a structured tree (data and tracking, segments, platforms, releases, external factors), ranks the hypotheses and gives the queries to run.

diagnose-metric-drop
PromptProduct metrics

Analyse a conversion funnel

Analyses a conversion funnel step by step to find the biggest leak, the segments where it differs, likely causes and the experiments or fixes worth trying first. For PMs and growth teams.

analyze-conversion-funnel
PromptProduct metrics

Build a growth experiment backlog

Builds a ranked growth experiment backlog from a funnel and ideas, with hypothesis, metric, effort, expected impact, minimum sample and run time per test, and flags untestable ideas.

build-experiment-backlog
PromptProduct metrics

Define an activation metric

Finds a product's activation moment from usage and retention data, defines an activation metric with an action, threshold and time window, and plans how to validate it.

define-activation-metric