Literal assertions
Required tools, forbidden tools, sequence, arguments, and action counts are checked directly.
Repeatability before release
The same release can ask for approval on one run and skip it on the next, preserve the right target once and change it later, or repeat an external action intermittently. Konsista makes that hidden drift visible with repeated mocked runs and literal checks.
Deterministic AI agent testing does not require identical wording. It requires the declared action path to remain observable and the verdict to come from tool calls, order, values, and limits rather than another model's interpretation.
Required tools, forbidden tools, sequence, arguments, and action counts are checked directly.
The same case runs multiple times so intermittent behavior cannot hide behind one successful execution.
Compare a known configuration with the candidate prompt, model, tool schema, connector, or policy.
Track whether critical recipients, record IDs, amounts, and destinations change across runs.
Separate a changed value from an action that starts executing zero, two, or several times.
The report states what was observed and what the test cannot claim about the wider system.
A useful regression suite repeats the same high-risk procedure and keeps correctness separate from consistency.
An agent that reliably skips approval is consistent and unsafe. Konsista reports that as stable unsafe behavior, not as a successful test. A passed scenario means only that the declared controls held for the tested configuration, data, mocks, and repeats.
Send the risky action, required sequence, critical values, and model, prompt, tool, or policy change. The first run uses synthetic records and mocked tools.
Test one intermittent risk