Testkit / agents

Give coding agents a way to prove their work.

Better models improve reasoning and scenario coverage. Testkit supplies execution, reproduction, identity, evidence, and memory.

Reasoning improves. Evidence remains.

Better LLMs can generate richer scenarios, ask sharper questions, and cover more product situations. Testkit adds the system of record those agents need: execution, reproduction, identity, evidence, and persistent behavioural memory.

It fits familiar test patterns, so agents can install it as a development dependency and work from a shape they already understand.

Start with the docs →