Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Use when writing up and releasing the results of a psychological audit of an AI/ML personnel assessment — producing a precise, comprehensive technical report for testing professionals AND a layperson-friendly summary for those the predictions affect, establishing the auditor's standards and credibility in the report, and deciding on public release. Triggers: "write the AI audit report", "release the bias audit results", "dual-audience audit report", "should we publish the audit", "auditor credib
| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-01 | ✗→✓ | ▲ Improved | -16% | 0% |
| case-02 | ✗→✓ | ▲ Improved | -3% | 0% |
| case-03 | ✗→✓ | ▲ Improved | -8% | 0% |
| case-04 | ✗→✓ | ▲ Improved | -8% | 0% |
| case-13 | ✗→✓ | ▲ Improved | 41% | 0% |
An audit only creates value if its results are communicated to the right audiences and, where the public interest is at stake, released. Reporting is also where the auditor's own credibility is established or lost.
Present results in multiple formats to meet the needs of all relevant audiences. At minimum:
enough that another competent professional could evaluate and, ideally, reproduce the audit's reasoning. Mirror the structure of a technical-validation-report, but organized around the 12 components and the claims evaluated.
(e.g., candidates) — clear, accurate, non-technical, addressing the fairness concerns in the terms those audiences actually use (justice/transparency — see ai-fairness-lenses).
One format cannot serve both audiences; write both.
validity / utility / lack-of-bias (from ai-audit-planning).
ai-fairness-lenses), statedso conclusions are interpretable across disciplines.
with evidence, gaps, and access limitations encountered.
audits) to improve the model.
intersectional subgroups, etc.).
No auditor or audit is automatically credible. "This system has been audited and is therefore credible" deserves skepticism. So the report must let readers judge the audit itself by disclosing:
in due course across the components;
discretion (the existence and scope of withholding should itself be disclosed).
Access and documentation vary even within auditor type, so transparency about access is essential to interpreting the findings.
Unless there is a compelling, transparently stated reason not to, an audit whose results are in the public interest should be released. Organizations may choose otherwise, but doing so risks the credibility of the audit, the company that built the algorithm, and the auditors. Public-facing, transparent, open audits are especially warranted when a system has outsized societal impact.
Normalize routine auditing. Treat regular internal and external auditing as a public good that raises the probability algorithmic systems in general are valid, valuable, and fair — and that builds public trust. Formative auditing folded into development, with complete documentation, can diminish or even preclude the need for post-hoc audits.
can't evaluate it).
ai-audit-planning (audience & release policy set up front) · ai-fairness-lenses · all model/stakeholder/meta audit skills · technical-validation-report (structure parallel)
Source: Landers & Behrend (2023), "Designing an Effective Psychological Audit" — multiple-format reporting, releasing results in the public interest, normalizing routine auditing, and auditor credibility.
Other measured skills in the registry, with their headline benchmark lift.