The four heights
The same task, four distances: today's deadline, the next reviewer, the stuck moment, the pattern.
Execute — do the immediate task
+We ran system tests last quarter and now need a compact performance summary for operations. Create…
Execute — do the immediate task
+We ran system tests last quarter and now need a compact performance summary for operations. Create a sheet that lists each test scenario in column A, expected throughput or metric in B, actual measured value in C, delta in D, and pass/fail in E. Add a column F with short notes where anomalies occurred and freeze the header row. Put a line chart comparing expected versus actual for the top five scenarios on a dashboard sheet.
Pasted it? When the reply comes back, push once: ask it to sharpen the weakest part. — Did this prompt help?
Improve — make it easier to accept
+Before I send the performance deck to the engineering manager, make it easy to see where the system…
Improve — make it easier to accept
+Before I send the performance deck to the engineering manager, make it easy to see where the system misses requirements. Put the three worst-performing scenarios up top, show the root-cause note adjacent to the delta, and flag any measurements taken under nonstandard conditions. Add a one-line recommendation for each flagged scenario that the manager can forward to the test lead.
Pasted it? When the reply comes back, push once: ask it to sharpen the weakest part. — Did this prompt help?
Decide — diagnose the stuck moment
+During the nightly load test throughput dropped 40% below spec for the payment gateway. The team…
Decide — diagnose the stuck moment
+A test run shows throughput 40% below expected under load
During the nightly load test throughput dropped 40% below spec for the payment gateway. The team involved is ops who scheduled the test and the performance engineer who wrote the harness. I don't know whether the harness misreported metrics or the gateway genuinely degraded. I can't surface this to the product owner until I'm confident it's not a measurement error. What's the likeliest diagnosis and the best quick check I should run in the spreadsheet to validate the metric?
Pasted it? When the reply comes back, push once: ask it to sharpen the weakest part. — Did this prompt help?
Become — change the pattern
+We waste cycles rerunning tests because the team treats any anomaly as a hard failure. Propose a…
Become — change the pattern
+We repeatedly accept flaky test results and re-test later
We waste cycles rerunning tests because the team treats any anomaly as a hard failure. Propose a habit to reduce needless reruns: one mandatory validation step to run when a result is out of tolerance, one place to log the validation outcome in the spreadsheet, and one rule for when to stop re-testing and escalate.
Pasted it? When the reply comes back, push once: ask it to sharpen the weakest part. — Did this prompt help?
Where the evidence lives
Who was seen doing this, and what people really ask.
Software tasks in the LLOS Work Atlas come from evidence, never a feature list: careers attested to do the work, real job descriptions, and the questions people actually ask (with their view counts). Facets — feature, workflow, troubleshoot, administer, deploy, scale — are open metadata: the work decides, not a taxonomy.
Copyright © LLOS.ai · 2026 — original pedagogy, voice, and design — all rights reserved.
The rest of the map
Same library, five ways in.