mirror of
https://github.com/magnus919/agent-skills.git
synced 2026-09-11 19:47:12 +03:00
* fix: SkillOpt epoch 1 for AI operating economics Promote cold-load entry points, quick-start reference routing, the minimum decision-record contract, and trigger-oriented progressive disclosure. Signed-off-by: Magnus Hedemark <magnus919@pm.me> * fix: SkillOpt epoch 2 for AI operating economics Add review-depth selection, evidence-to-disposition guidance, and scenario-led routing across adjacent skills. Signed-off-by: Magnus Hedemark <magnus919@pm.me> * fix: SkillOpt epoch 3 for AI operating economics Expose a minimum claim ledger and explicit closure conditions for every bounded disposition. Signed-off-by: Magnus Hedemark <magnus919@pm.me> * fix: resolve SkillOpt review consistency findings Align entry-point paths, canonical step routing, claim-ledger fields, and triage disposition wording. Signed-off-by: Magnus Hedemark <magnus919@pm.me> * fix: resolve final SkillOpt disposition wording Keep review-depth outputs inside the canonical disposition set and distinguish supported claims from permitted language. Signed-off-by: Magnus Hedemark <magnus919@pm.me> * fix: complete SkillOpt routing correction Route triage through the outcome-map step and identify the evidence-classification step explicitly. Signed-off-by: Magnus Hedemark <magnus919@pm.me> * fix: complete AI economics review template Add the minimum decision-record fields required by the optimized skill routing contract. Signed-off-by: Magnus Hedemark <magnus919@pm.me> --------- Signed-off-by: Magnus Hedemark <magnus919@pm.me>
3.5 KiB
3.5 KiB
AI Initiative Evidence Record
Decision header
- Initiative:
- Workflow:
- Population and scope:
- Intervention mode: assist / recommend / route / execute / replace
- Decision sought: scale / constrain / redesign / hold / retire / exception
- Accountable decision owner:
- Review date or trigger:
- Record version:
Value hypothesis
For [population] doing [workflow], [intervention] will change [outcome] by [direction/range] without exceeding [countermetric boundary], at [full cost boundary], compared with [baseline], over [period].
- Hypothesis status: supported / weakened / refuted / unresolved
- Expected benefit:
- Enabled capacity:
- Operational benefit:
- Realized economic or mission benefit:
- Benefit realization mechanism:
- Human-control boundary:
- Authority proposed for next slice:
Evidence comparison
- Baseline:
- Comparison design:
- Treatment period:
- Comparison period:
- Inclusion and exclusion rules:
- Known selection effects:
- Concurrent changes:
- Quality measurement limitations:
- Statistical or causal analysis owner:
| Claim | Evidence class | Source and scope | What it supports | What it does not support | Open challenge | Permitted interpretation |
|---|---|---|---|---|---|---|
| observed / causal / inferred / vendor-reported / asserted / normative |
Outcome and countermetrics
| Metric | Type | Definition and denominator | Baseline | Observed | Target/boundary | Owner | Evidence source |
|---|---|---|---|---|---|---|---|
| primary / leading / countermetric / adoption |
Segment review
| Slice | Adoption or exposure | Outcome | Quality/countermetric | New burden or benefit | Decision implication |
|---|---|---|---|---|---|
Slices not available and why:
Cost boundary
Billing truth
- Provider/infrastructure source:
- Billing period and version:
- Reconciliation status:
Allocated cost
- Allocation target:
- Shared-cost rule:
- Allocation owner:
Economic cost
- Meaningful unit:
- Model/inference:
- Tools and APIs:
- Retrieval/storage/networking:
- Human review and exception handling:
- Incremental capacity:
- Engineering, evaluation, observability, governance, training, support, and exit:
- Fixed, variable, step-function, avoided, transferred, and uncertain costs:
- Calculation and allocation method:
- Range or sensitivity:
Findings and gaps
Supported findings
Unresolved or conflicting findings
Missing evidence
| Gap | Why it matters | Owner | Next evidence | Due date or trigger |
|---|---|---|---|---|
Governance evidence packet
- Intended use and risk tier:
- System/model/prompt/policy/tool/provider/version inventory:
- Acceptable-use, refusal, escalation, and human-oversight rules:
- Pre-deployment evaluation and release threshold:
- Third-party/provider assessment and contractual evidence:
- Incident, override, and near-miss record:
- Change/revalidation trigger:
- Retention, dependency, leakage, user-impact, and decommissioning plan:
Decision and controls
- Disposition:
- Scope of approval:
- Authority limit:
- Budget or quota limit:
- Human review or escalation rule:
- Stop trigger:
- Rollback, containment, or retirement path:
- Exception approver, if applicable:
- Revisit condition:
Learning closure
- Expected versus observed outcome:
- Expected versus observed cost:
- Countermetric and subgroup result:
- Incidents, overrides, or near misses:
- Changes since prior record:
- Hypothesis update:
- Follow-up artifact or owner: