Research

AI Governance & Multi-Model Decision Research

Original research on multi-model agreement, disagreement, human authorization, decision traceability, evaluator reliability, and governed AI workflow performance.

Evidence standard: no fabricated statistics. Quantitative benchmarks publish only when measured corpora are available.

Program

How to use this library

Start with the reproducible evaluator methodology (program anchor). Review the multi-model agreement benchmark design for the latest benchmark framing. Use governance guides for operational context, and the Citation Kit on each report for attribution.

LATEST RESEARCH

Research library

The Refusal Gap

How AI Models Differ in What They Will and Will Not Answer

Methodology published; quantitative benchmark results forthcoming

When Humans Override AI

What Governed Review Reveals About Model Reliability

Study design published; measured Decision Ledger results forthcoming

The Economics of Governed AI

Measuring Review, Escalation, and Decision Cycle Time

Measurement design published; workflow telemetry results forthcoming

BENCHMARKS

BENCHMARKS

The Refusal Gap

How AI Models Differ in What They Will and Will Not Answer

Methodology published; quantitative benchmark results forthcoming

The Economics of Governed AI

Measuring Review, Escalation, and Decision Cycle Time

Measurement design published; workflow telemetry results forthcoming

METHODOLOGY

METHODOLOGY

DECISION GOVERNANCE

DECISION GOVERNANCE

HUMAN OVERSIGHT

HUMAN OVERSIGHT

When Humans Override AI

What Governed Review Reveals About Model Reliability

Study design published; measured Decision Ledger results forthcoming

DATA & EVALUATION

DATA & EVALUATION

Related governance

Authority pages

See governed AI execution in a live workflow

Review how SmartSolo coordinates multiple AI models, routes human authorization, and preserves the decision record.