What a Gaper AI-agent claim means, and what it does not.
We separate verified results, client-reported observations, external benchmarks, and modeled planning outputs. The label stays with the claim so a reader can see the evidence standard without reading fine print.
Gaper’s evidence methodology is a claim-labeling and review system that identifies the source, calculation, limitations, and approval status behind a statement about an AI-agent workflow.
Four labels, four different evidence standards.
The qualifier remains adjacent to the claim in page copy, cards, metadata, social assets, and exported material.
A result calculated from retained source data using a documented definition, time period, calculation version, and review record.
- May be used for
- A client-approved, scope-specific measured outcome with its methodology and limitations stated near the claim.
- May not be used for
- A generalized promise, an unreviewed dashboard screenshot, or an outcome without an identifiable baseline and source trail.
- Dated baseline and post-launch source exports
- Metric definition, calculation, and version
- Attribution method and material confounders
- Client approval and named reviewer signoff
Render “Verified result” next to the metric, with a link or expandable methodology note. State the client scope and measurement window.
A statement made by a client or operator that Gaper has not independently verified against the underlying operational data.
- May be used for
- A clearly attributed qualitative observation or a client-supplied metric when the source and wording are approved.
- May not be used for
- An independently verified outcome, a benchmark, or a promise of comparable results.
- Written client approval
- Exact approved wording and attribution
- Date and operating context
- Disclosure that Gaper did not independently verify the underlying data
Render “Client-reported” immediately before the statement. Do not remove the qualifier in cards, metadata, social copy, or comparison tables.
A statistic, research finding, standard, or market observation published by an external source and used for context rather than proof of a Gaper implementation.
- May be used for
- Market context, workflow-selection rationale, or a comparison point with a direct, current source link.
- May not be used for
- A claimed Gaper result, a forecast for a client, or a substitute for measuring the client’s baseline.
- Primary or authoritative source where available
- Publication date and URL
- Scope, methodology, and population checked for relevance
- A statement that the benchmark does not predict a client outcome
Render “External benchmark” with a direct source link and describe the population or context. Never present it as a customer result.
An illustrative calculation or workflow scenario derived from stated assumptions. It is useful for prioritization, not evidence of a realized outcome.
- May be used for
- Cost or ROI planning, scenario comparison, implementation reference designs, and hypothesis setting before measurement.
- May not be used for
- A customer case study result, guaranteed savings, a quote, delivery commitment, or compliance conclusion.
- Visible input assumptions
- Formula or calculation method
- Disclosure of excluded benefits and costs
- A prompt to replace assumptions with a dated baseline and post-launch data
Render “Modeled” adjacent to every output and repeat the disclosure in exported, shared, or downloadable versions.
Planning math stays planning math.
Calculator results use user-entered assumptions and transparent formulas. They are directional planning outputs, not pricing, legal, compliance, accounting, medical, staffing, or revenue advice. A result becomes publishable only after a baseline, post-launch data, an attribution method, evidence retention, and the required client and expert approvals are complete.
Use the strongest source the claim can support.
Higher levels can support stronger, narrower claims. A lower-level source cannot be promoted through confident copy.
A trigger creates a required review action.
Evidence pages are not “set and forget.” Claims pause when the evidence or operating condition changes materially.
- 01
New page or material claim
Assign an actual reviewer, select the claim label, confirm the source hierarchy, and record disclosure text before publication.
- 02
New client outcome or quote
Collect client approval, baseline and post-launch source data, metric definition, calculation version, limitations, and reviewer signoff before changing the label to verified or client-reported.
- 03
Pricing, product scope, policy, integration, or regulatory change
Review linked public ranges, controls, language, and source citations before the page is updated or reused in a new campaign.
- 04
At least every six months
Recheck benchmark dates, broken links, calculation defaults, named-reviewer status, and whether a modeled scenario has been mistakenly reused as a result.
- 05
Quality or safety incident
Pause affected claims, document the incident and corrective action, revalidate the evaluation set and controls, then restore content only after review.
What this method cannot promise.
AI-agent performance depends on workflow variance, data quality, access, policy design, evaluation coverage, human review, rollout, and ongoing operation. It does not transfer automatically between organizations.
A lower average handling time can conceal a worse exception path. Measure rework, corrections, escalation, safety, and customer or operator impact alongside speed.
No page should imply that a modeled output is a client result, that a benchmark predicts a result, or that a client-reported observation was independently verified.
Privacy, security, clinical, legal, accounting, employment, and compliance decisions require the client’s qualified owners and advisors. This methodology is not professional advice in those domains.
Evidence questions.
What is the difference between verified and client-reported?+
Why label a result as modeled?+
Can an external benchmark prove ROI for our workflow?+
How often does Gaper update evidence pages?+
Build an evidence plan for one workflow
Define the claim, baseline, control boundary, evidence pack, and reviewer before launch.