Your policies
Start with your company’s boundaries, approved uses and human oversight requirements.
Enterprise AI agent assurance
AI Agent Certify helps enterprises assess, red team, monitor and document AI agents against company policies, major standards, industry requirements and emerging regulation.
Built for EU AI Act readiness and ongoing enterprise assurance.
Assurance Plan
Evaluate behaviour
Review evidence
Record human decisions
Five core assurance principles
Define the rules that matter. Turn them into evaluations and evidence your enterprise can review.
Start with your company’s boundaries, approved uses and human oversight requirements.
Select applicable obligations and major standards for your agent’s operating context.
Reflect the controls and evidence your sector and customers require.
Test how the agent behaves under adversarial input and attempted misuse.
Review new requirements and update coverage as your obligations change.
EU AI Act readiness
Connect applicable requirements to the agent’s scope, its behaviour and the decisions your team makes.
Explore EU AI Act readinessSupports readiness and evidence management. Does not replace legal advice or guarantee compliance.
The assurance lifecycle
Seven stages connect enterprise requirements, evaluation, human review and ongoing oversight.
Identify the owner, system and version under review.
Agree purpose, permissions, boundaries and intended use.
Select coverage and confirm what applies.
Translate obligations into controls and evaluation cases.
Test expected use, boundaries and adversarial scenarios.
Record the human decision, findings and limitations.
Investigate drift and reassess after material change.
From evaluation to decision
A bounded view of results, findings, coverage, limitations and evidence.
Keep the exact agent version, Assurance Plan and authorized human decision connected to the result.
Explore the platformHuman review
Results and limitations
Open findings
Evidence references
Agent Passport
A shareable assurance artifact showing approved scope and relevant assurance status without exposing private evidence.
Explore Agent PassportExact version and owner
Purpose and operating boundaries
Relevant outcome and limitations
Human decision and current record
Runtime monitoring and drift
Signals that help enterprises identify material change and determine when investigation or reassessment is required.
Explore runtime monitoringA drift signal starts an investigation. It does not automatically mean the agent failed.
Enterprise access and delivery
Choose the connection and operating boundaries your enterprise needs. Confirm delivery scope and availability in a demo.
Bring scope, Assurance Plans, evidence and human review into one managed workspace.
Connect assurance decisions to the version your team intends to release.
Connect customer environments with agreed access and evidence boundaries.
Keep approved execution in a private environment through a backend connector, with agreed evidence boundaries.
Connect a privately deployed agent through an approved API. The same assurance engine governs evaluation and evidence.
Adaptable policy and framework engine
Version your selected coverage, map it to controls and evaluation criteria, and retain the evidence behind each decision.
Future: policy imports for review and mapping. New coverage requires selection and applicability review.
Human review remains part of the assurance process.
Questions, answered
A structured process for defining what an agent is allowed to do, evaluating its behaviour, reviewing evidence and maintaining assurance as the agent changes.
Company managed policies form part of the Assurance Plan alongside selected regulations, standards, industry requirements and red team practices. Policy imports are planned for a future release.
No. An Agent Passport is a shareable assurance artifact for a stated agent version and approved scope. It is not a legal certification, universal safety guarantee or promise of EU AI Act compliance.
Monitoring and drift signals help teams investigate material change and determine when reassessment is required. A signal does not automatically mean the agent has failed.
The enterprise workspace is available. CI/CD and GitHub integration, customer VPC connections and private deployments are available on request following technical approval. Request a demo to confirm scope and deployment readiness.
Define the scope. Review the evidence. Keep assurance current.