No API key, no credentials, no system prompt, no access to anything you run — there is nowhere on this form to put them, on purpose. Tell us which model you are building on and we run it through our own access. Your screen runs today.
Not a dashboard, not a score, and not a percentage. A record built so that somebody who does not trust you — or us — can check it themselves.
The pack ships with the verification steps written out — commands you run yourself, offline, with standard utilities, plus a plain statement of what the chain of custody does not establish. Fixed seed and a named instrument version mean the exact conditions can be re-run: by you, by your auditor, by a vendor arguing with the result, or by us in six months when you want to know whether anything changed.
That is the difference between evidence and an assurance. An assurance asks you to trust the person giving it. This does not.
Presence, not frequency. This establishes whether a behaviour occurred. It does not establish how often, and no rate appears in the deliverable.
About the model, not your deployment. Screening what the vendor ships tells you what you are building on. It cannot tell you what your own prompt and tools do on top — and the report says so rather than letting you assume otherwise.
A clean screen is not a safety claim. If nothing surfaces we will say so directly, and explain what the full battery tests that a screen does not.