Behavioral Verification Is Not Optional
for AI Agents in Regulated Industries.
Here Is Why TTIC Is Testing It
as a Certification Requirement.
Why Existing Frameworks Fail for
Agentic AI in Regulated Industries
Every compliance framework in healthcare was designed for a deterministic world. You write a policy, configure a control, conduct an audit. The system does what you told it to do. Documentation of intent maps roughly to documentation of behavior because the system does not deviate from its configuration.
AI agents break that assumption at a fundamental level. Their behavior is probabilistic, context-dependent, and changes over time without anyone modifying the configuration. SOC 2 does not measure this. HIPAA technical safeguards do not measure this. Most of ISO 27001 does not measure this. They measure whether controls were documented and in place during an audit window. That is a different question from whether the system actually behaved correctly under the full range of conditions it will encounter in production, including adversarial ones.
This is the gap that TTIC was built to address. And it is the gap that made Praxen immediately compelling when Steve Wilson brought it to our attention.
What Enterprise Clinical AI Governance
Actually Requires
Most vendors present clinical AI governance as two layers: model accuracy and basic vendor documentation. Enterprise clinical AI governance requires eleven. The realization that Praxen and Darwin together covered all eleven layers is what inspired the TTIC certification pathway.

Not as a product decision, but as a governance completeness realization. Praxen addresses the technical and behavioral layers through pre-deployment and ongoing verification. Darwin governs the clinical transaction layers continuously after deployment. Together they cover the full stack.
This layer reflects TTIC's independent analysis of regulatory and compliance considerations, including first-principles alignment with FDA's published cybersecurity and quality system frameworks. No FDA review or endorsement of this framework has been sought or obtained.
What Agent Behavior Verification Is
and Why Regulated Industries Need It
Agent Behavior Verification is a new and important discipline in AI security. It is the practice of probing an AI system's actual behavior against its intended behavior under adversarial conditions, before deployment and on an ongoing basis after deployment. Praxen is the open-source reference implementation of ABV, built by Steve Wilson and released under Apache 2.0 by Exabeam.
Praxen reflects and extends Exabeam's established leadership in behavior intelligence and agent security. Where Exabeam has long led the field in detecting anomalous behavior by humans and non-human identities at runtime, Praxen extends that leadership upstream into pre-deployment verification, closing the gap between what an AI agent is configured to do and what it actually does before it ever reaches production. It is a natural and significant extension of a leadership position Exabeam has built over years.
For regulated industries, the stakes are distinct from the general enterprise context. In healthcare, behavioral deviation is not a productivity loss. It is a patient safety event, a liability exposure, and an institutional accountability failure. A clinical AI system that behaves correctly 98% of the time and deviates 2% of the time in ways that affect clinical decisions is not a software quality problem. It is a governance failure.
Praxen's RAISE framework scores behavioral maturity across six categories: domain limitation, knowledge base balance, zero trust implementation, supply chain governance, adversarial testing, and continuous monitoring. Each maps directly to governance requirements that regulated industries face. Darwin is engineered to complement this with deterministic, auditable clinical transaction scoring that enterprises can rely on by design rather than by assumption.
Understanding RAISE: A Clinical AI Governance Perspective
By Sherri Douville CEO, Medigram · Founder & Chair, TTIC · Co-Chair, Trust and Identity Subgroup, IEEE/UL 2933 · Series Editor, Taylor & Francis · Brian Yam Chair, Pro Sports Track, TTIC · COO, Somnology
Each category below includes framing for:
Click on the button for your title to see how this applies to you.
Click on the button for your title to see how this applies to you.
Click on the button for your title to see how this applies to you.
Click on the button for your title to see how this applies to you.
Click on the button for your title to see how this applies to you.
Click on the button for your title to see how this applies to you.
What We Learned Running It
Medigram accessed Praxen through its early access program before public launch and ran it against Darwin, the governed clinical AI platform, across three successive assessment cycles. TTIC evaluated the results in its capacity as the clinical governance body with the expertise to interpret what the findings mean in a clinical context.
The third run, using Praxen 0.8.0, produced the following result:
What the three-run cycle taught TTIC was more valuable than the final score. It taught us what the categories actually mean in a clinical context and how to interpret findings in terms of clinical accountability rather than just security posture. That interpretive layer is precisely why TTIC's early access certification program requires Praxen as Layer 1. The tool produces rigorous findings. Understanding what those findings mean for a governed clinical platform requires the expertise TTIC brings. A RAISE score issued without that interpretive layer is a number. Within the TTIC certification process it becomes a clinical accountability document.
The Stadium
and the Field
The best analogy for what Praxen and Darwin together certify is a stadium and a field. The infrastructure is certified and ready. The behavioral verification has been conducted. The governance architecture is in place. But the teams still must coach and the players still have to play the game.
That is the honest scope of what technical and behavioral certification provides. It verifies that the infrastructure meets the standard required for governed clinical play. It does not play the game for you. The clinical workflows, the physician oversight structures, the accountability decisions made by health system leaders — those are the game. The TTIC certification early access program is testing whether this infrastructure is consistently ready for it.
How Praxen Complements
Runtime Governance
Praxen is a pre-deployment and ongoing behavioral verification instrument. Darwin's governed transaction architecture ensures that every clinical decision made by the system after deployment is captured as an immutable governance record, with identity, scope, and audit trail intact, as a structural condition of the transaction completing. One verifies before. The other governs continuously after.
TTIC's early access certification program requires both because clinical AI governance is not a point-in-time event. It is a continuous accountability structure. Praxen gives the pre-deployment evidence. Darwin gives the runtime evidence. Together they produce the documentation chain that health system procurement, D&O underwriters, and the institutional accountability structures of regulated industries require.
What the Early Access
Certification Program Requires
TTIC is currently testing its certification pathway through an early access program. The pathway is structured in two layers. Layer 1 is Praxen RAISE behavioral verification. Layer 2 is Darwin TIPPSS governance screening across six clinical dimensions.
TIPPSS screening functions as a required input to TTIC certification; certification itself is issued by TTIC's governance review, not by TIPPSS scoring alone.
| Tier | Scope | RAISE Threshold | TIPPSS Requirement |
|---|---|---|---|
| Tier 1 | Direct clinical data contact | 4.0 or above | TIPPSS Level 3 · All six dimensions · Documented adversarial testing |
| Tier 2 | Operational & infrastructure vendors | 3.0 or above | TIPPSS screening · Three dimensions |
The first step for any vendor or health system interested in the early access program is defining the governance committee with authority to sign off on the work remit for a Praxen run. This is not an engineering decision. It is a clinical accountability decision. The remit defines scope, the accountable parties for findings, and the remediation pathway. Without that structure, the Praxen run produces findings with no governance pathway to resolution.
Praxen is free and open source. You can begin exploring it today. For information about participating in the TTIC certification early access program, contact Sherri Douville directly.
Author & Organizations
The launch also had meaningful reach. In the first few days post, the Praxen launch generated more than 20 placements across cybersecurity, AI, and enterprise technology media, including coverage in North America, Asia-Pacific, the UK, Europe, and Africa. Medigram as a technical demonstration under TTIC was the only clinical AI governance organization with a practitioner account published on launch day.
Praxen is free and open source. TTIC's certification early access program is open to health systems and vendors ready to demonstrate governed clinical AI.
Discussed on The New CISO (Exabeam), Episode 150 · Episode highlights
Choose TTIC as a preferred source in eligible Google experiences. Prefer TTIC in Google
- Published by
- Trustworthy Technology & Innovation Consortium (TTIC)
- Author
- By Sherri Douville, Founder & Chair, Trustworthy Technology & Innovation Consortium (TTIC)
- Originally published
- Last updated
Cite this resource
Sherri Douville. “Behavioral Verification Is Not Optional for AI Agents in Regulated Industries.” Trustworthy Technology & Innovation Consortium (TTIC), 2026. https://trustworthytechnologyinnovation.com/blog-praxen-behavioral-verification.