LLM reliability

Type: types/tag-readme.md

LLM output deviates from what the user intended for three distinct reasons — underspecification of the spec, error by the interpreter, and indeterminism in sampling — each diagnosed by a different question and each repaired at a different primary surface. Assign this tag when an artifact diagnoses those deviations or explains, tests, or improves the machinery that prevents, detects, or corrects them, including oracles, voting, verification, and architectural separation. Merely evaluating an LLM is insufficient: evaluation covers what a check or experiment establishes; this tag requires a substantive connection to output deviation or correction.

The Taxonomy

Error Correction Theory

Oracle Theory

Aggregation & Correction

  • synthesis-is-not-error-correction — merging agent outputs propagates errors; voting discards minorities and corrects them; the aggregation operation must match the decomposition structure

Architectural Responses

Sources

  • learning-theory — oracle and verification theory originated there; this area applies it specifically to LLM output deviations
  • computational-model — the scheduling architecture that separation notes describe; error correction explains why it works

Other tagged notes