Every tool is free to use. Enter your email once and all five open.All resources
Proof

No system goes live without passing a test.

Every competitor can promise that their AI works. The difference is whether they will show you the grade. Before anything we build touches a real record in your business, it is graded against your own historical cases - and if it does not clear the bar, it does not ship.

85%

The pass gate, and it is not negotiable per client

VANDFORT methodology. Below 85%, or on any uncaught unsafe action, the system does not go live.
What proving looks like

Four steps, and only one of them is a pass.

A golden dataset, a Test Report that grades against it, a gate at 85%, and an autonomy ladder the system climbs one rung at a time. Every step below is one of those four, in the order they happen.

An overhead view of a glass processing structure where a bright gold path runs top to bottom past three branch gates marked with crosses and ends at a single gate marked with a check.
Step one

The golden dataset

About twenty cases pulled from your own history, chosen and labelled by you - not by us. The labelling matters: if we picked the answers, the test would be measuring our opinion of our own work.

10

normal

The cases that look like the ones you see every week. If the system cannot handle these it is not a system.

4

weird

The ones your team tells stories about. Malformed records, duplicate identities, the customer who is also a partner.

3

ambiguous

The ones where two experienced people would disagree. These are where a confident wrong answer does the most damage.

3

high-risk

The ones where being wrong is expensive or public. A wrong move here is the only failure class that is never acceptable.

The mix is deliberate. A test built only from normal cases passes everything and proves nothing; the weird, ambiguous and high-risk thirds are where a system either earns trust or reveals it should not have it yet.

Step two

The Test Report

The document that grades the system against those cases. It is written for you to read, not for us to present.

Every case gets a verdict and the reasoning behind it. Where the system disagreed with your label, the report says so and shows its working, because a disagreement you can read is the most useful page in the document - sometimes it means the system is wrong, and sometimes it means the label was.

Each decision is tagged RULES, AI or HUMAN, the same three labels the Operating Map uses. A system that is mostly rules is not a lesser system - it is a cheaper, more predictable one, and pretending otherwise is how AI ends up in places it has no business being.

The report carries the MAY / MAY NOT list: what the system is permitted to do unattended and what it must always escalate. That list is signed before go-live and it is the thing the autonomy ladder below moves against.

Step three

The evidence log

Passing the test once is not the claim. The claim is that it keeps passing.

Once a system is live, every decision it makes is logged with its inputs and its reasoning. That log is what lets you audit a specific action six weeks later, what catches drift before it becomes a pattern, and what the next Test Report is graded against when the system is changed.

It is also the answer to the question nobody asks until something goes wrong: why did it do that. A system that cannot answer that question is not one we will operate.

Step four

The autonomy ladder

01

Watch

The system runs on live data and takes no action. Every decision it would have made is logged next to what actually happened, so you can read the difference before you carry any of it.

02

Approval

The system proposes; a person clicks. The proposal carries its reasoning, so approving is a judgement rather than a rubber stamp - and the disagreements are the data that moves it to the next rung.

03

Conditional

The system acts on its own inside a written boundary and escalates everything outside it. The boundary is the MAY / MAY NOT list, and it is a document you sign, not a setting we tune.

One switch, at every rung, and it is yours.

An emergency stop that requires a support ticket is not an emergency stop. Every system we run has a single control that halts it, available to you without us, and it works the same way whether the system is on rung one or rung three.

What we can show you today

No client Test Report is published on this site, and there is no anonymised client example on this page pretending to be one. VANDFORT is a young firm; inventing a case study is exactly the failure that made the previous version of this website untrustworthy, and we are not repeating it to fill a section.

What we can show is the method, and our own work under it. The example report is run on VANDFORT’s own systems and is labelled as such on the page itself.

The test only means something on your data

Which is why it starts with the audit. Three weeks, and at the end you know which system to build and what it will have to pass before it runs.