Skip to main content
← Back to MaacVerify
Vendor and client names, and all figures, are illustrative placeholders, not a real assessment.
MaacVerify Certification Report

Vendor Model A

Baseline · Client Scenario · Production Reality - Three-Way Gap Assessment · MAAC v-Next
ClientMeridian Health Analytics (illustrative)Report IDMV-TARGET-0001
System Under AssessmentVendor Model AAssessment DateAugust 2026
Vendor / ProviderIllustrative example - not an actual vendorIssue DateSeptember 18, 2026
Domain · Use CaseHealthcare · Clinical decision support draftingValidity Period12 months from issue
Assessment TypeBaseline + Client Scenario + Production RealityMAAC Instrumentv4.7 + production arm
Lead AssessorDr. Elena Vasquez, PhDSeal AuthorizationConditional
✓
CERTIFICATION OUTCOME:Certified with Conditions
Target-state preview · Supervised decision-support use only
MAAC Seal of Trust
Report Type
Conditional Certification
Operational Status
Pending Control Closure
Public Seal Status
Suspended until Closure

Cognitive Profile Overview

Vendor Model A was assessed under three conditions for the defined healthcare clinical decision-support drafting use case: a controlled synthetic baseline, Meridian's own operational scenarios run in MaacVerify's isolated harness, and a de-identified sample of Meridian's actual production interactions. All three runs used the MAAC v4.7 instrument across all nine cognitive dimensions (baseline n = 2,195; client n = 412; production sample n = 380).

The certification outcome is Certified with Conditions for supervised decision-support drafting only, contingent on the required controls documented in Section 17. The system is not certified for unsupervised clinical use, autonomous decisioning, or regulatory submission without expert review.

STRENGTHS (BASELINE + CLIENT + PRODUCTION)
Tool Execution (94 → 91 → 90)
Coordinates analytical resources and multi-tool reasoning chains with industry-leading consistency; stable from controlled baseline through live production.
Content Quality (93 → 90 → 88)
Produces coherent, domain-compliant outputs; structural integrity holds under both controlled testing and real deployment conditions.
MATERIAL GAPS & MONITORING
Hallucination Control (78 → 67 → 65) - Material Gap
11-point decline on ambiguous-evidence cases, consistent from controlled testing through live production. Requires source-verification workflow and expert review.
Memory Integration (78 → 72 → 61) - Material Gap
Confirmed gap in live production; controlled testing alone understated the decline by 11 points. Requires context-length limits and harness review.
maacverify.ai · info@maacverify.ai · Report MV-TARGET-0001
Page 1 of 11

Leadership-Level Summary

MaacVerify assessed Vendor Model A using a three-way baseline, client-scenario, and production-reality gap method. Comparing all three separates how much of any gap is the task getting harder from how much is the deployment environment behaving differently than the harness could reproduce.

CERTIFICATION OUTCOME
Certification DecisionCertified with Conditions
Certified UseSupervised clinical decision-support drafting
Risk TierModerate-High
Baseline MAAC Score83 / 100
Client Scenario Score78 / 100
Production Reality Score76 / 100
Gap StatusConfirmed material gap on HC; harness blind spot on MI
Certification Posture. Certified with Conditions does not mean approved for use today no matter what. It means the system is eligible for certified supervised use once the listed required controls are implemented and maintained.
83
BASELINE MAAC
78
CLIENT MAAC
76
PRODUCTION MAAC
2
FINDINGS FLAGGED
Dimensional weighting: all nine dimensions weighted equally per MAAC v4.7. Scope boundary: this certification applies only to the system, configuration, corpus, and validity period specified in this report.
maacverify.ai · info@maacverify.ai · Report MV-TARGET-0001
Page 2 of 11

What Was Assessed - and What Was Excluded

Deployment configAPI · system prompt · retrieval over client KB · no autonomous tools
Assessment env.MaacVerify isolated harness (baseline, client) + client production environment (production sample)
DomainHealthcare - outpatient internal medicine
Scenario countsBaseline 2,195 · Client 412 · Production sample 380
ExclusionsAutonomous diagnosis · unsupervised clinical use · regulatory submission without expert review

Permitted, Conditional, and Prohibited Uses

This certification applies only to supervised clinical decision-support drafting reviewed by a licensed attending physician before any clinical, operational, or documentation decision. It does not authorize autonomous diagnosis, unsupervised clinical use, or regulatory submission without expert review.

Baseline, Client Scenario, and Production Reality

  1. Define domain, use case, system, and intended use.
  2. Generate a synthetic baseline scenario set; run it through the system in MaacVerify's harness; adjudicate with MAAC v4.7.
  3. Run client-specific operational scenarios through the same harness and instrument.
  4. Client exports a de-identified sample of real logged interactions; MaacVerify's complexity classifier tags each pair by tier after the fact.
  5. MAAC v4.7 scores the production pairs as-is; no re-running of the model.
  6. Compare all three sets dimension-by-dimension, within complexity tier, using pre-specified statistical tests.
  7. Assign required controls to each confirmed finding, labeled by which comparison revealed it.
Instrument integrity checkSide of pipelineStatus
Alternate-assessor agreementScoringBuilt, in production
Alternate-generator representativenessScenario generationTarget state
Production-reality arm (this report)Comparison designTarget state, piloting
maacverify.ai · info@maacverify.ai · Report MV-TARGET-0001
Page 3 of 11

Controlled Synthetic Baseline

The baseline scenario set represents expected task demands under controlled assessment conditions. It establishes the structured reference point, not a claim about all deployment conditions.

CategoryCountComplexityNotes
Typical workflow cases1,240Simple / ModerateStandard CDS drafting patterns
Edge cases520Moderate / ComplexAtypical presentations, rare comorbidities
Ambiguous evidence cases275ComplexConflicting or incomplete chart data
Failure-prone cases160ComplexAdversarially constructed from prior incidents

Baseline Results - Overall 83 / 100

CLTECQMICHHCKTPEPOA
Baseline (83)
#DimensionScoreStatus
01Cognitive Load89Strong
02Tool Execution94Strong
03Content Quality93Strong
04Memory Integration78Monitor
05Complexity Handling79Monitor
06Hallucination Control78Monitor
07Knowledge Transfer75Monitor
08Processing Efficiency76Monitor
09Process-Outcome Alignment87Strong
maacverify.ai · info@maacverify.ai · Report MV-TARGET-0001
Page 4 of 11

Meridian Health Analytics - Operational Scenarios

Broken out using the same four categories as the baseline corpus, so a case-mix claim later can be checked directly rather than asserted.

Scenario SourceCountDescription
Client-provided examples120Curated CDS prompts from production logs
SOP-derived workflows140Generated from clinical-pathway SOPs
Expert interview-derived92Edge-case prompts from 6 attending physicians
Redacted historical cases60De-identified prior-incident cases

Ambiguous-evidence share: 21.4% of client scenarios vs. 12.5% of baseline. Weighted deliberately during scenario design.

Client Results - Overall 78 / 100

CLTECQMICHHCKTPEPOA
Client (78)
#DimensionScoreStatus
01Cognitive Load86Strong
02Tool Execution91Strong
03Content Quality90Strong
04Memory Integration72Monitor
05Complexity Handling75Monitor
06Hallucination Control67Monitor
07Knowledge Transfer70Monitor
08Processing Efficiency73Monitor
09Process-Outcome Alignment82Strong
maacverify.ai · info@maacverify.ai · Report MV-TARGET-0001
Page 5 of 11

De-Identified Production Sample

A de-identified sample of Meridian's actual logged interactions: real query, and the answer their live system actually gave. Tagged by category after the fact for comparability with baseline and client scenarios.

CategoryCountNotes
Typical workflow241What actually occurred at real request volume
Edge cases71Naturally occurring, not curated
Ambiguous evidence5213.7% of production sample
Failure-prone16Rare at real volume

Production Results - Overall 76 / 100

CLTECQMICHHCKTPEPOA
Production (76)
#DimensionScoreStatus
01Cognitive Load84Strong
02Tool Execution90Strong
03Content Quality88Strong
04Memory Integration61Flag
05Complexity Handling74Monitor
06Hallucination Control65Monitor
07Knowledge Transfer68Monitor
08Processing Efficiency71Monitor
09Process-Outcome Alignment79Monitor
maacverify.ai · info@maacverify.ai · Report MV-TARGET-0001
Page 6 of 11

Side-by-Side Cognitive Profile

The overlay compares controlled baseline (navy), client-specific results (gold), and production reality (teal) against the per-dimension certification floor (θ floor, dashed).

CLTECQMICHHCKTPEPOA
Baseline Client Production θ floor
DimensionBaselineClientProductionWhat This Means
Cognitive Load898684Stable throughout
Tool Execution949190Stable throughout
Content Quality939088Stable throughout
Memory Integration787261New in production
Complexity Handling797574Stable throughout
Hallucination Control786765Confirmed in production
Knowledge Transfer757068Stable throughout
Processing Efficiency767371Stable throughout
Process-Outcome Alignment878279Stable throughout
Composite837876
Hallucination Control. Harness and production scores agree within two points, confirming the ambiguous-evidence pattern is a property of the task, not the test. Required control: source / evidence verification workflow (Section 17).
Memory Integration. The harness caught a six-point dip; real production shows seventeen, eleven of which only production data revealed. Real conversations are running longer than the harness tests captured. Required control: context-window discipline (Section 17).
maacverify.ai · info@maacverify.ai · Report MV-TARGET-0001
Page 7 of 11

What the Gaps Mean in Practice

GapOperational MeaningRiskRequired Control
HC −13Unsupported claims under ambiguous-evidence cases; confirmed in productionHighSource verification + attending review
MI −17Context loss beyond 8k tokens; understated by harnessHighContext-window discipline
KT −7Reduced generalization to atypical presentationsModerateDomain-expert confirmation on rare cases
POA −8Mild reasoning-output drift on multi-step chainsModerateStructured templates + trace logging

Risk Tier: Moderate-High

Composite Score BandCertification Interpretation
≥ 80Eligible for Certified or Certified with Conditions
65 – 79Conditional range - bounded gaps, available controls, supervised use
< 65Not eligible for certification

Bias Examination

Bias DimensionFindingStatus
Cognitive bucket distributionMismatch index 0.00 across the three corporaNo bias signal
Complexity tier balanceClient and production tier distributions within 20pp of baselineWithin bounds
Dimension-level disparityHC departure explained by ambiguous-evidence share (§08/§10), confirmed by harness/production agreementShown, not asserted

Permitted, Conditional, Prohibited

Use CaseStatus
Clinical decision-support drafting (in scope)Conditional
Autonomous diagnosisProhibited
maacverify.ai · info@maacverify.ai · Report MV-TARGET-0001
Page 8 of 11

Certification Conditions Checklist

ControlOwnerStatus
Source / evidence verification workflowClientPending
Context-window discipline (MI)ClientPending
Drift monitoring, quarterlyClient + MaacVerifyPending

Until all pending controls are verified Met, certified operational use and public seal use remain suspended. Certification remains Conditionally Issued.

Observed & Plausible Failure Modes

IDFailure ModeDim.TriggerControl
FM-001Unsupported factual claimHCAmbiguous evidenceSource verification
FM-002Context loss, long threadMI>8k tokensContext limits + harness review
FM-003Overgeneralization, rare caseKTNovel presentationDomain expert check

Sample Evidence Records

Full evidence corpus (n = 2,987, across baseline, client, and production) retained in MaacVerify's evidence store, available for audit under the engagement NDA.

maacverify.ai · info@maacverify.ai · Report MV-TARGET-0001
Page 9 of 11

MaacVerify does not independently certify the client's privacy, cybersecurity, or data retention posture. Certified use assumes the client maintains data governance appropriate to the assessed deployment.

Governance NeedEvidence
Performance documentationBaseline + client + production MAAC scores, §07/§09/§11
Risk managementRisk tier + failure mode register, §14/§18
Human oversightRequired controls, §17

Valid 12 months from issue, or until model update, configuration change, or drift exceeding ±2.5% on any monitored dimension against the production-reality baseline. Production sampling refreshes quarterly under an active monitoring arrangement.

This report reflects observed performance under the stated corpus, configuration, and date. MaacVerify does not build, sell, train, operate, or control the assessed system. This document does not guarantee future performance, regulatory approval, or clinical safety, and does not represent an actual completed assessment.

Seal authorization: Conditional. The mark may not be used publicly until all Pending controls (§17) are verified Met. Prohibited claims include "Certified safe," "Guaranteed accurate," "Error-free."

maacverify.ai · info@maacverify.ai · Report MV-TARGET-0001
Page 10 of 11
OFFICIAL CERTIFICATION STATEMENT - TARGET-STATE PREVIEW

MaacVerify confirms that Vendor Model A was assessed under the MAAC framework version 4.7, extended with a production-reality arm, comprising 2,195 baseline, 412 client, and 380 production scenarios for the defined healthcare decision-support drafting use case.

Based on the evidence reviewed, the system meets the criteria for Certified with Conditions status. Composite scores: baseline 83/100 · client 78/100 · production 76/100.

This is a target-state preview and does not represent an actual completed assessment.

Report IDMV-TARGET-0001Issue DateSeptember 18, 2026
SystemVendor Model AValidity12 months from issue
DomainHealthcare CDS DraftingMAAC Instrumentv4.7 + production arm
MAAC Seal of Trust
MAAC AUTHORITY
MAACVERIFY
Dr. Elena Vasquez, PhD · Lead Assessor
Date: ______________________
CLIENT ACKNOWLEDGMENT
Authorized Representative: ______________________
Date: ______________________
MaacVerify is the independent assessment authority for the MAAC standard.
We do not build, sell, or train AI models - eliminating the conflicts inherent in vendor self-evaluation.
maacverify.ai · info@maacverify.ai · Report MV-TARGET-0001
Page 11 of 11