Long-Horizon Dependability Beyond Single Sessions
Determine how the Claude Code architecture—comprising the agentic query loop, context construction and compaction, subagent delegation, and session persistence—can continue to support long-horizon dependability when autonomous work spans beyond a single session.
References
How the architecture documented in \Cref{sec:arch,sec:turn,sec:context,sec:subagent,sec:persist} (whose primary units are the turn, the session, and the sub-agent) continues to support long-horizon dependability as autonomous work extends beyond a single session is an open question.
This would allow us to test whether covert strategies remain internally consistent across memory boundaries and handoffs, whether goal-task conflicts compound when strategic behavior is distributed, and whether patterns observed here, such as the role of instrumental goals or the reasoning-action gap.
Whether a human-AI collaboration can sustain those obligations over the multi-year lifetime of a solver, rather than over a single task, remains an open question.