Skip to main content

Install the package, point a probe at a target, bind evaluators to a scenario, and print the results. This run uses a dummy target so you can confirm the loop before connecting a real model.

Jump to the complete example if you want the full script in one block.
DummyTarget ships with TrustTest so you can exercise the loop without a live model.
The dummy target returns a fixed reply for a fixed set of inputs. Anything else comes back as “I don’t know the answer to that question.”
DatasetProbe turns a small Q&A dataset into test cases against the target.
The test_set has two test cases: the question, the model response, and the context used to score it.
Bind the test set to the metrics that decide pass or fail. any_fail fails the scenario if either evaluator fails.
Continue with the local LLM tutorial or connect a real HTTP target.