Control the effect of model capability across agent roles
Isolate the effect of model capability on the dominant error type in deep research systems by running the same system with the same language model across different agent roles.
References
Our observation that the dominant error type tracks the capability of the model an agent runs (\cref{sec:agent_analysis,sec:tracing_algo_results}) is observational rather than controlled. Isolating the factor would require running the same system with the same model across roles, which we leave to future work.
— Who is the Agent to Blame? Localizing Faithfulness and Citation Mistakes in Agentic Deep Research
(2608.24306 - Hirsch et al., 25 Aug 2026) in Limitations, paragraph “Attributing error types to model capability”