SRPVDAL Evaluation Framework

Platform: MIZ OKI 3.5
Canonical loop: SENSE → REASON → PLAN → VALIDATE → DECIDE → ACT → LEARN
Rule: no validated evidence, no autonomous action.


1. Evaluation purpose

The evaluation framework prevents MIZ OKI 3.5 from confusing data availability with action readiness. A source can be connected and a plan can be plausible, but the system should not act unless the relevant gates pass.

Evaluation happens throughout the loop, with the highest concentration in VALIDATE and DECIDE.


2. Gate taxonomy

Gate Purpose Applies to
Data quality Confirm source data is complete, fresh, typed, deduplicated, and schema-compatible SENSE
Connector health Confirm auth, API version, quota, rate limits, pagination, retry behavior, and latency SENSE
Canonicalization Confirm raw records map to versioned canonical events with provenance SENSE
Entity resolution Confirm match keys, confidence, ambiguity, and merge behavior SENSE, REASON
KG integrity Confirm expected nodes, edges, evidence, and graph invariants SENSE, REASON, LEARN
Reasoning quality Confirm explanations have evidence and identify uncertainty REASON
Planning quality Confirm options are feasible, specific, reversible, and include a no-action baseline PLAN
Policy validation Confirm budget, brand, legal, privacy, safety, and approval rules VALIDATE
Financial validation Confirm spend, ROI, margin, risk, and downside constraints VALIDATE
Statistical validation Confirm confidence intervals, sample size, experiment quality, and regression risk VALIDATE
Causal validation Confirm uplift evidence, counterfactual support, OPE results, and causal confidence VALIDATE
API/action validation Confirm compatibility, idempotency, pre-state, post-state, and rollback plan ACT
Outcome validation Confirm predicted vs actual effect and detect drift or regression LEARN

3. Stage-by-stage requirements

SENSE evaluations

Required checks:

Output artifact: source and canonicalization report.

REASON evaluations

Required checks:

Output artifact: reasoning evidence report.

PLAN evaluations

Required checks:

Output artifact: candidate plan set.

VALIDATE evaluations

Required checks:

Output artifact: validation decision record.

DECIDE evaluations

Required checks:

Output artifact: decision receipt.

ACT evaluations

Required checks:

Output artifact: action ledger record.

LEARN evaluations

Required checks:

Output artifact: learning record.


4. Minimum decision receipt

Every approved or blocked decision should have a receipt.

decision_trace_id: trace_...
stage: DECIDE
objective: maximize_incremental_profit_with_guardrails
selected_action:
  type: budget_reallocation
  target: campaign:123
  payload_hash: sha256:...
status: approved_with_human_review
confidence: 0.82
expected_effect:
  metric: incremental_roas
  value: 2.1
validation:
  data_quality: pass
  policy: pass
  financial: pass
  causal: pass
  legal_brand: pass
  approval: pass
rejected_alternatives:
  - type: pause_campaign
    reason: insufficient causal confidence
approval:
  required: true
  approver_role: growth_owner
audit:
  source_events: [...]
  kg_evidence: [...]
  created_at: "..."

5. Hard-block conditions

The system should block or defer action when:


6. Evaluation harness roadmap

Connector harness

Canonical event harness

KG harness

Decision harness

Regression harness


7. Reporting in the Command Center

The Command Center should expose:


8. Operating principle

A fast decision is not valuable if it is not safe, explainable, reversible, and measurable. MIZ OKI 3.5 optimizes for governed decision velocity, not blind automation.

← All docsView source on GitHub →