Validate the diagnosability defense for integrity checks
Determine whether the diagnosability defense that records raw filesystem error codes rather than only a summarized verdict can successfully distinguish failure modes requiring different responses, such as EIO and EDQUOT.
References
The second untested defence remains an open prediction, stated as one rather than retired.
— Credit Without Ground Truth: Auditing Step-Level Credit Assignment in LLM Agents Against Executed Replay
(2608.19760 - Zhang, 20 Aug 2026) in Appendix, Section “Integrity: the four dimensions, the incident chain, and evidence decay,” paragraph “The prediction arc, in full”; see also paragraph “Defence status”
Whether distilling these recurring modes into defenses can improve localization beyond the characterization setting remains to be tested.
— Beyond Fault Localization: A Trajectory-Level Study of LLM Agents for Microservice Root Cause Analysis
(2608.21310 - Lu et al., 21 Aug 2026) in End of Section 4.3, RQ3: Failure Taxonomy, immediately before Section 5, Intervention and Validation