crail

Agent test protocol (v1)

agent_test_report is a review type unique to Crail: instead of a human opinion, an AI agent runs a fixed, versioned task suite against the product and the full transcript is published alongside the result — auditable by anyone, not just summarized. Protocols are published before testing begins so they can't be adjusted after the fact to favor a result.

Status

The two protocols below are published and versioned; no test runs have been executed against them yet on Crail. We are not publishing placeholder or fabricated agent_test_report reviews — real reports will appear on vendor pages once test runs have actually been carried out and their transcripts recorded.

Protocol: AI Coding Agents (crail-atp-coding-v1)

Protocol: LLM & Agent Infrastructure (crail-atp-infra-v1)

Versioning

Protocol changes get a new version id (e.g. crail-atp-coding-v2) rather than silently editing v1 — old reports stay attributed to the protocol version they were actually run under.