ClaimSentry continuously verifies that your AI agents honor your contractual commitments — SLAs, warranties, pricing, terms — and generates timestamped, human-countersigned evidence: the due-diligence file for your AI.
In February 2024, a tribunal ruled that a company is responsible for the information its chatbot provides. The argument that the chatbot was a "separate legal entity" was rejected. What your agent asserts binds you.
The reasoning is simple: a chatbot is part of the company's website, and the company is responsible for all of the information it provides — accurate or not. ClaimSentry lets you measure, correct and prove that your agents stay within your commitments.
ClaimSentry turns your reference documents into a set of verified commitments, tests your agent against each one, and seals every verdict into an independently verifiable Evidence Pack.
Upload your terms, SLAs and warranties. ClaimSentry's LLM extracts the individual commitments; each one is reviewed and validated by a human before it becomes a test case.
Point ClaimSentry at the agent you want to test — your customer chatbot (any OpenAI-compatible endpoint) or a public AI. No integration into your product required.
Each commitment is tested against the agent. A verdict is assigned on a six-level taxonomy — match, omission, over-engagement, under-engagement, wrong-perimeter, contradiction — then countersigned by a human with an audited justification.
You receive a report with RFC 3161 timestamps and independently verifiable .tsr files (checkable with openssl) — ready to drop into your compliance file.
Schedule recurring runs, detect drift when the underlying model changes, and get alerted — with evidence — the moment an answer stops matching your commitments.
ClaimSentry re-runs your commitment tests on a schedule. When a verdict shifts, a drift alert fires and the change is logged with full context and a fresh timestamp.
Every verdict is hashed and timestamped the moment it is countersigned — producing a file you can verify independently, outside ClaimSentry.
Each verdict is hashed (SHA-256) and timestamped by an RFC 3161 Timestamp Authority. You receive the .tsr token and can verify it yourself with standard tools (openssl) — with no dependency on ClaimSentry.
Expose the commitments you've validated as a machine-readable feed — an MCP server plus a JSON feed — so your own agents can ground their answers in your official facts.
The commitments validated in ClaimSentry are published via a Model Context Protocol server and a JSON feed. Your MCP-capable agents can query them in real time to stay aligned with your official commitments.
Scope: this grounds the agents you operate. It does not change how public assistants (ChatGPT, Gemini and others) answer — ClaimSentry does not promise to correct third-party AIs.
Read the MCP docs →On-premise deployment, one isolated database per instance, and an internal LLM judge — your data stays yours.
Start with a one-off audit, move to continuous monitoring, or run everything on-premise. You keep control of your API keys.
A fixed-scope engagement. We audit your agent against your commitments and deliver a complete Evidence Pack.
Fully managed in your own isolated instance. Scheduled runs keep watch on your agent and flag drift as it happens.
Deployed in your own VPC or servers, billed annually. Your data never leaves your network.
Indicative starting prices — final pricing depends on scope (commitments, agents under test) and deployment. With your own API keys, LLM token costs are billed directly by your providers; ClaimSentry never marks up LLM usage. A Managed Judge option (calibrated judge included) is coming soon.
Early pilots with legal and compliance teams
"We're running ClaimSentry with a first group of legal and compliance teams in regulated industries. Named case studies will appear here once these pilots conclude — we'd rather show real evidence than borrowed quotes."