Files
magnus919_agent-skills/ai-operating-economics/templates/ai-initiative-evidence-record.md
T
Magnus HedemarkandGitHub 1291f9576e fix: SkillOpt optimize AI operating economics (#369)
* fix: SkillOpt epoch 1 for AI operating economics

Promote cold-load entry points, quick-start reference routing, the minimum decision-record contract, and trigger-oriented progressive disclosure.

Signed-off-by: Magnus Hedemark <magnus919@pm.me>

* fix: SkillOpt epoch 2 for AI operating economics

Add review-depth selection, evidence-to-disposition guidance, and scenario-led routing across adjacent skills.

Signed-off-by: Magnus Hedemark <magnus919@pm.me>

* fix: SkillOpt epoch 3 for AI operating economics

Expose a minimum claim ledger and explicit closure conditions for every bounded disposition.

Signed-off-by: Magnus Hedemark <magnus919@pm.me>

* fix: resolve SkillOpt review consistency findings

Align entry-point paths, canonical step routing, claim-ledger fields, and triage disposition wording.

Signed-off-by: Magnus Hedemark <magnus919@pm.me>

* fix: resolve final SkillOpt disposition wording

Keep review-depth outputs inside the canonical disposition set and distinguish supported claims from permitted language.

Signed-off-by: Magnus Hedemark <magnus919@pm.me>

* fix: complete SkillOpt routing correction

Route triage through the outcome-map step and identify the evidence-classification step explicitly.

Signed-off-by: Magnus Hedemark <magnus919@pm.me>

* fix: complete AI economics review template

Add the minimum decision-record fields required by the optimized skill routing contract.

Signed-off-by: Magnus Hedemark <magnus919@pm.me>

---------

Signed-off-by: Magnus Hedemark <magnus919@pm.me>
2026-08-21 16:21:23 -04:00

3.5 KiB

AI Initiative Evidence Record

Decision header

  • Initiative:
  • Workflow:
  • Population and scope:
  • Intervention mode: assist / recommend / route / execute / replace
  • Decision sought: scale / constrain / redesign / hold / retire / exception
  • Accountable decision owner:
  • Review date or trigger:
  • Record version:

Value hypothesis

For [population] doing [workflow], [intervention] will change [outcome] by [direction/range] without exceeding [countermetric boundary], at [full cost boundary], compared with [baseline], over [period].

  • Hypothesis status: supported / weakened / refuted / unresolved
  • Expected benefit:
  • Enabled capacity:
  • Operational benefit:
  • Realized economic or mission benefit:
  • Benefit realization mechanism:
  • Human-control boundary:
  • Authority proposed for next slice:

Evidence comparison

  • Baseline:
  • Comparison design:
  • Treatment period:
  • Comparison period:
  • Inclusion and exclusion rules:
  • Known selection effects:
  • Concurrent changes:
  • Quality measurement limitations:
  • Statistical or causal analysis owner:
Claim Evidence class Source and scope What it supports What it does not support Open challenge Permitted interpretation
observed / causal / inferred / vendor-reported / asserted / normative

Outcome and countermetrics

Metric Type Definition and denominator Baseline Observed Target/boundary Owner Evidence source
primary / leading / countermetric / adoption

Segment review

Slice Adoption or exposure Outcome Quality/countermetric New burden or benefit Decision implication

Slices not available and why:

Cost boundary

Billing truth

  • Provider/infrastructure source:
  • Billing period and version:
  • Reconciliation status:

Allocated cost

  • Allocation target:
  • Shared-cost rule:
  • Allocation owner:

Economic cost

  • Meaningful unit:
  • Model/inference:
  • Tools and APIs:
  • Retrieval/storage/networking:
  • Human review and exception handling:
  • Incremental capacity:
  • Engineering, evaluation, observability, governance, training, support, and exit:
  • Fixed, variable, step-function, avoided, transferred, and uncertain costs:
  • Calculation and allocation method:
  • Range or sensitivity:

Findings and gaps

Supported findings

Unresolved or conflicting findings

Missing evidence

Gap Why it matters Owner Next evidence Due date or trigger

Governance evidence packet

  • Intended use and risk tier:
  • System/model/prompt/policy/tool/provider/version inventory:
  • Acceptable-use, refusal, escalation, and human-oversight rules:
  • Pre-deployment evaluation and release threshold:
  • Third-party/provider assessment and contractual evidence:
  • Incident, override, and near-miss record:
  • Change/revalidation trigger:
  • Retention, dependency, leakage, user-impact, and decommissioning plan:

Decision and controls

  • Disposition:
  • Scope of approval:
  • Authority limit:
  • Budget or quota limit:
  • Human review or escalation rule:
  • Stop trigger:
  • Rollback, containment, or retirement path:
  • Exception approver, if applicable:
  • Revisit condition:

Learning closure

  • Expected versus observed outcome:
  • Expected versus observed cost:
  • Countermetric and subgroup result:
  • Incidents, overrides, or near misses:
  • Changes since prior record:
  • Hypothesis update:
  • Follow-up artifact or owner: