Log experiment results

Log experiment results in Notion — with the four heights of help laid out: do it now, make it easier for the next person to accept, work out the right move when you are stuck, and learn the pattern so it stops coming back.

4prompt heights
Open it in the interactive atlas →

The four heights

The same task, four distances: today's deadline, the next reviewer, the stuck moment, the pattern.

Execute — do the immediate task

+
Record the A/B landing test results from last week: test name, variant A and B conversion rates,…
Record the A/B landing test results from last week: test name, variant A and B conversion rates, sample sizes, start and end dates, primary metric lift, statistical significance, and the implementation decision. Link each result to the Growth experiments database and tag it with 'retention' or 'acquisition'. Verify the metric math and mark the record complete for Friday’s review.

Improve — make it easier to accept

+
Before I add last week’s experiment to the experiments log, make the outcome readable for the…
Before I add last week’s experiment to the experiments log, make the outcome readable for the product and marketing leads: put the headline result first (e.g., Variant B +12% signups), show the raw counts and confidence interval right below, call out any secondary metrics harmed, and flag what would make a stakeholder hesitate to build it out.

Decide — diagnose the stuck moment

+
We ran an A/B test where signups increased but revenue per user dropped 6%. The PM wants to ship…

Signups rose but revenue per user dipped slightly

We ran an A/B test where signups increased but revenue per user dropped 6%. The PM wants to ship the winning signup flow; marketing wants the growth. I’m unsure which metric to prioritize in the record and afraid I’ll bias the log to please the PM. What’s the most likely correct diagnosis of the trade-off, and what precise wording and tags should I use so future reviewers can see both effects without taking sides?

Become — change the pattern

+
Over time our experiments log became a mess: different metrics, inconsistent tags, unclear…

Experiment records are inconsistent and hard to compare

Over time our experiments log became a mess: different metrics, inconsistent tags, unclear decisions. That costs me hours reconciling results before roadmap meetings. Where do we repeatedly lose value, what single template or habit would make 80% of entries usable by anyone, and how do I get product and marketing to adopt it without policing every entry?

Next to this one

Other docs and databases work people do in Notion.

Every task here came from the work, not from a feature list — which is why the prompts name what you want done and never the button that does it. The tool changes; the work does not.
Copyright © LLOS.ai · 2026 — original pedagogy, voice, and design — all rights reserved.

The rest of the map

Same library, five ways in.