# Human in the loop

`kaal:entity:human-in-the-loop`

**Status.** derived

This node is assembled mechanically from the 5 claims that carry the concept tag `human-in-the-loop`. It is a roster of what the corpus says under this term. It is **not** an adjudicated definition: no single statement here has been ruled canonical, and no first-appearance call has been made. Read the claims and judge for yourself.

## Every claim under this term

5 claims across 4 works, 2024 to 2025.

**2024**

- [4796714-017](https://wulfkaal.github.io/claims/4796714-017) [failure/argued] *(failure mode)* -- Using human judgment to uncover unconscious bias in AI can perpetuate the very biases it is meant to remove, because human reviewers carry their own implicit biases and may lack the expertise to identify bias in complex AI systems.
  > While human judgment is integral to risk management and bias mitigation, it inherently carries its own biases. Relying on human judgment to uncover unconscious biases in AI may inadvertently perpetuate these biases rather than eliminate them.
  Wulf A. Kaal, AI Governance (2024). SSRN: https://ssrn.com/abstract=4796714
- [4855607-016](https://wulfkaal.github.io/claims/4855607-016) [failure/evidenced] *(failure mode)* -- There is a trade off in RLHF between the agent imitating human advice and learning autonomously, and human guidance that is too specific will prevent the agent from discovering novel optimal strategies.
  > there is a trade-off between the extent to which the agent should imitate human advice versus learning autonomously. Overspecific human guidance can hinder the agent's ability to discover novel optimal strategies.
  Wulf A. Kaal, How AI Models are Optimized Through Web3 Governance (2024). SSRN: https://ssrn.com/abstract=4855607

**2025**

- [5095633-021](https://wulfkaal.github.io/claims/5095633-021) [failure/argued] *(failure mode)* -- Human-in-the-loop annotation, including under ethical labor models, imposes financial and time costs large enough to slow the pace at which AI models can be upgraded.
  > The reliance on human annotators, even with companies striving for ethical labor practices like CloudFactory, involves significant costs, both financial and in terms of time, which can slow down the pace of AI model upgrades.
  Wulf A. Kaal, Artificial Intelligence The Final Frontier (2025). SSRN: https://ssrn.com/abstract=5095633
- [5541658-010](https://wulfkaal.github.io/claims/5541658-010) [design/evidenced] -- The Shenzhen courts illustrate a workable three step human-in-the-loop design for generative judicial AI: the judge decides first, the LLM generates the supporting reasoning, and the judge then revises that output to finalize the judgment.
  > A notable case study from Shenzhen, China, illustrates a three-step interaction pattern: judges make initial decisions, LLMs generate reasoning based on these decisions, and judges revise the output to finalize judgments.
  Wulf A. Kaal, Morgan A. Gray, The Evolving Role of Artificial Intelligence in Law (2025). SSRN: https://ssrn.com/abstract=5541658
- [5541658-016](https://wulfkaal.github.io/claims/5541658-016) [failure/argued] *(failure mode)* -- Keeping judges ultimately accountable through a human-in-the-loop review of AI generated reasoning, as practiced in Shenzhen, does not fully resolve the accountability problem in AI assisted adjudication.
  > Shenzhen case study illustrates that judges retain ultimate accountability, revising AI-generated reasoning to ensure accurate judgments, but this human-in-the-loop approach does not fully resolve the issue.
  Wulf A. Kaal, Morgan A. Gray, The Evolving Role of Artificial Intelligence in Law (2025). SSRN: https://ssrn.com/abstract=5541658

## Verify

Every claim above resolves to a record carrying a verbatim source quote, the sha256 of the source PDF, and a preformatted citation. Nothing here asks to be taken on trust.

    curl -s https://wulfkaal.github.io/entities/human-in-the-loop.md | sha256sum

**Canonical form.** This markdown file is the canonical hashed representation of this entity node. Its sha256 is the content hash.
