# Trustworthiness

`kaal:entity:trustworthiness`

**Status.** derived

This node is assembled mechanically from the 3 claims that carry the concept tag `trustworthiness`. It is a roster of what the corpus says under this term. It is **not** an adjudicated definition: no single statement here has been ruled canonical, and no first-appearance call has been made. Read the claims and judge for yourself.

## Every claim under this term

3 claims across 2 works, 2025 to 2026.

**2025**

- [5541658-011](https://wulfkaal.github.io/claims/5541658-011) [normative/argued] -- Accuracy alone is insufficient for legal AI: a model must also be explainable before its outputs can be trusted in judicial settings.
  > Accuracy alone is insufficient without explainability.
  Wulf A. Kaal, Morgan A. Gray, The Evolving Role of Artificial Intelligence in Law (2025). SSRN: https://ssrn.com/abstract=5541658
- [5541658-012](https://wulfkaal.github.io/claims/5541658-012) [failure/argued] *(failure mode)* -- Post hoc explainability techniques do not by themselves establish trustworthiness; the explanations they produce must additionally be verified against human knowledge.
  > However, post-hoc explainers still need to be verified for trustworthiness, in that the explanations comport with human knowledge.
  Wulf A. Kaal, Morgan A. Gray, The Evolving Role of Artificial Intelligence in Law (2025). SSRN: https://ssrn.com/abstract=5541658

**2026**

- [6244278-027](https://wulfkaal.github.io/claims/6244278-027) [mechanism/argued] -- When reputation is both the system's aggregation weight and the agent's optimization target, equilibrium behavior includes consistency, collaborative integrity, and systemic stewardship, so the agent learns not merely to perform well but to be trustworthy.
  > when reputation is both the system's aggregation weight and the agent's optimization target, the equilibrium behavior includes properties, such as consistency, collaborative integrity, systemic stewardship, that transcend narrow task optimization.
  Wulf A. Kaal, AI's Mother's Instinct Engineered Consequence Emergent Ethics and the Institutional Trajectory Toward Agentic Alignment (2026). SSRN: https://ssrn.com/abstract=6244278

## Verify

Every claim above resolves to a record carrying a verbatim source quote, the sha256 of the source PDF, and a preformatted citation. Nothing here asks to be taken on trust.

    curl -s https://wulfkaal.github.io/entities/trustworthiness.md | sha256sum

**Canonical form.** This markdown file is the canonical hashed representation of this entity node. Its sha256 is the content hash.
