Sopact is a technology based social enterprise committed to helping organizations measure impact by directly involving their stakeholders.
Copyright 2015-2026 © sopact. All rights reserved.
Describe your program and Sopact Sense maps the results framework — strategic objective to intermediate results to sub-IRs — grading each with its baseline indicator and flagging the results that don't ladder up.
In short: One prompt now builds a funder-ready results framework from nothing but what a program publicly states. It names the decision the reader must make, arranges the program as a hierarchy — one Strategic Objective, a handful of Intermediate Results, Sub-IRs beneath them — grades every result by evidence, flags the results that don't ladder up and the indicators with no baseline, and ends with a prioritized plan. Below, we run it on a real public page and walk through what each part of the prompt produces.
A results framework is the hierarchy USAID and similar funders ask for: it shows how a program's lower-level results add up to its top aim — one Strategic Objective, several Intermediate Results that must each occur to reach it, and Sub-IRs beneath those, every level carrying a baseline-and-target indicator. Most results frameworks fail two ways: a result that doesn't ladder up to anything above it, or an indicator with a target but no baseline to judge it against.
The prompt below turns that logic into a complete audit. It reads only what your program states publicly, grades every result Green, Amber, or Red, flags the orphans and the missing baselines, and tells you the one result to fix this quarter. We ran it on The Lantern Network's public mentoring-program page — the sections that follow show the prompt for each part, then what came back.
Every grade depends on who is reading. A board sees "87% placed" as a headline; a renewing funder asks what evidence sits behind each result, and whether every result actually rolls up to the objective. So the prompt's first instruction is a decision frame: before building anything, state in one line who would use this results framework and for what decision. For Lantern, that came back as a corporate sponsor deciding whether to renew its grant. Every judgment below is made from that reader's chair.
Here is the full prompt. Paste it whole, swap in your program name and source, and Sense produces all four parts in one pass:
Build a Results Framework for [PROGRAM NAME] using only what the program publicly states at [SOURCE — URL or pasted program description]. Before building, state in one line who would use this framework and for what decision — then make every judgment from that reader's perspective. PART 1 — Map the hierarchy: one Strategic Objective, 3–5 Intermediate Results that must each occur to reach it, and the Sub-IRs beneath them. Give every level a baseline-and-target indicator. Flag any result that doesn't ladder up to the level above, and any indicator with no baseline. Color every element: GREEN = specific AND evidenced, with a baseline; AMBER = stated but vague, or a target with no baseline; RED = missing, doesn't ladder up, or exists only as [INFERRED]. Tag anything not explicitly stated as [INFERRED]. Include a legend, program name, source URL, and date. PART 2 — One row per result: Result | Level (SO / IR / Sub-IR) | Grade | Baseline & Target (or "none") | one measurable indicator that would test it, phrased so the program could actually collect it (who is measured, what changes, by when). PART 3 — List every orphan result and every AMBER and RED element, ranked by how much the framework depends on it. For each: why it is weak in one sentence, what claim collapses if it fails, and whether the fix is a program-design problem or a measurement gap. PART 4 — The top 3–5 fixes, in priority order for the decision named above. For each: current language (or "missing") → proposed rewrite, plus the single data collection step that would move it toward green. CLOSING SUMMARY — 3–4 sentences: overall strength of the results logic, the weakest link a skeptical funder would attack first, and the one action to take this quarter. RULES — Source fidelity is absolute: never invent program content; if the source does not say it, mark it RED or [INFERRED]. Every green grade must be traceable to specific source language.
Example source: https://www.lanternnetwork.org/mentoring-program. The rules do the heavy lifting — no invented content, every green traceable to a quote and a baseline. The hierarchy is the claim; the table in Part 2 is its evidence trail.
Map the hierarchy: one Strategic Objective, 3–5 Intermediate Results, and Sub-IRs beneath them, each with a baseline-and-target indicator. Flag any result that doesn't ladder up and any indicator with no baseline. Color every element green, amber, or red. Tag anything not explicitly stated as [INFERRED]. Include a legend, program name, source URL, and date.
Two disciplines make this different from a wish-list of outcomes. First, every result must ladder up: each Sub-IR rolls into an IR, each IR into the objective — anything that doesn't connect is flagged as an orphan. Second, no result is judged without a baseline: a target with nothing to measure against earns an amber at best. For Lantern, the pattern was immediate — the employment IR is concrete and baselined (288 mentees, 251 internships, 87% placed) while the sustained-outcome results past placement rest on three testimonials with no baseline at all.
The rubric is strict on purpose. Green means specific, evidenced, and baselined ("87% secured internships, jobs, or promotions"). Amber means stated but vague, or a target with no baseline. Red means missing, doesn't ladder up, or exists only as [INFERRED].
GRADE: green | employment IR | 87% placed, 251 internships — baselined, ladders up; amber | confidence sub-IR | claimed in three stories, no baseline; red | sustained-outcome IR | no baseline, and an orphan sub-IR that doesn't ladder up
For every result, produce one row: Result | Level (SO / IR / Sub-IR) | Grade | Baseline & Target (or "none") | one measurable indicator that would test it, phrased so the program could actually collect it — who is measured, what changes, by when.
This table is the evidence trail behind the hierarchy — one row per result, so nothing in the ladder floats free of a baseline. The indicator column is the practical payoff: each is phrased so the program could actually collect it. For Lantern's weakest result — the sustained-outcome IR — the baseline column reads "none: no follow-up past first placement," and the indicator that would fix it is "share of placed mentees still employed at 12 and 24 months, same mentees tracked by ID." That's not a critique; it's a work order.
List every orphan result and every amber and red element, ranked by how much the framework depends on it. For each: (a) why it is weak in one sentence, (b) what claim collapses if it fails, (c) whether the fix is a program-design problem or a measurement gap — these require different responses.
The ranking is by dependence, not level — the question is which weakness takes the most down with it. Lantern's number one wasn't its red impact claims; it was the missing baseline under the sustained-outcome IR: with no follow-up past placement, the whole top of the ladder rests on the 87% number alone. An orphan sub-IR — ongoing alumni support that connects to no IR — came second.
The design-versus-measurement tag matters just as much. A measurement gap means the result plausibly holds but has never been baselined — the fix is data collection. A design problem means a result genuinely connects to nothing — the fix is re-linking or removing it. Lantern's diagnosis came back mostly measurement gaps, plus one orphan to re-link.
Give the top 3–5 fixes, in priority order for the decision named above. For each: show the current language (or "missing") → a proposed rewrite, plus the single data collection step that would move it toward green.
Each fix is a before-and-after pair. Lantern's first: the sustained-outcome IR, currently a target with "no baseline," gets a stated baseline and target — "X% of placed mentees still employed at 12 months" — fed by a single data step: a short follow-up survey against a persistent participant ID. Fixes two through five follow the same shape: re-link the orphan alumni sub-IR to the employment IR, baseline the confidence sub-IR, and publish the denominator behind the 87%.
Close with 3–4 sentences: the overall strength of the results logic, the weakest link a skeptical funder would attack first, and the one action to take this quarter.
The summary is the executive read. Lantern's verdict: the framework is strong at its employment IR and unproven above it — the objective and its sustained-outcome results carry targets with no baselines, and one sub-IR ladders up to nothing. The one action this quarter: baseline the sustained-employment result with a 12-month follow-up against a persistent participant ID, so next year's renewal case rests on tracked results rather than three stories.
Take the prompt with you. The full prompt pack — the master prompt plus each part as a standalone prompt you can run separately — is available to download: Download the Results Framework prompt pack.
Every result must ladder up. The defining test of a results framework is that each sub-IR rolls into an IR, and each IR into the objective. Ask Sense to flag any result that doesn't connect — an orphan result is the most common reason a framework reads as incoherent.
No baseline, no IR. An intermediate result with a target but no baseline can't be judged. Ask Sense which results are missing a baseline and to suggest one the program could realistically collect.
Turn one result green per cycle. Fix the reddest result with the one indicator Sense suggests, collect it next cycle, and re-run the framework.
Tighten your program page while you're here. Once Sense has graded the framework, ask it to bring your public claims in line with your evidence:
Based on the grades above, suggest edits to my program page so its claims match the evidence. Flag every sentence that overstates what we can show, and rewrite it to be accurate and specific.
The same prompt works for a Theory of Change, Logic Model, or Logframe — swap the framework name and keep the parts, rubric, and rules unchanged.
A results framework is a hierarchy used in monitoring and evaluation that shows how a program's lower-level results add up to its top-level aim: one Strategic Objective, several Intermediate Results that must each occur to achieve it, and Sub-IRs beneath those — every level carrying a baseline-and-target indicator. It's favored by USAID and similar funders for showing the logic of “what leads to what.”
A results framework shows the hierarchy of results — how sub-results ladder up to intermediate results and to the strategic objective. A logframe is a 4×4 matrix that adds indicators, means of verification, and assumptions for each level. The results framework answers “what leads to what”; the logframe adds “how we'll measure and what must hold true.” Many funders ask for both.
Describe your program and ask the AI to map one strategic objective, three to five intermediate results, and their sub-IRs, each with a baseline-and-target indicator, then flag any result that doesn't ladder up and any missing baseline. In Sopact Sense this takes minutes and stays grounded only in what your program states, grading each element so the unlinked results and gaps are obvious.
Open Sopact Sense, paste your program description, and put it to work.
Try in Sopact