An AI evaluation improves after a prompt edit. A week later, a teammate cannot reproduce the comparison. The prompt file is available, but the retrieved note changed, a model alias may point somewhere else, and nobody recorded the adapter revision. There is an answer on disk without enough context to explain its origin. Create a run manifest that identifies the actual request inputs, configuration, source snapshot, and evaluator revision. Then assign that manifest a stable digest. The diges...