Papers
Topics
Authors
Recent
Search
2000 character limit reached

Agent-Level Behavior Models

Updated 16 July 2026
  • Agent-Level Behavior Models are defined as frameworks that capture individual agent decisions, policies, and trajectories, emphasizing local behavior over aggregate outcomes.
  • These models employ diverse representations—such as decision trees, Markov chains, and diffusion processes—to provide interpretable and predictive insights.
  • Applications range from traffic simulation and reinforcement learning to cognitive behavioral analysis, enabling targeted interventions and detailed performance evaluation.

Searching arXiv for recent and foundational papers on agent-level behavior models and closely related formulations. Using arXiv search to verify the cited papers and anchor the synthesis in the current literature. Agent-level behavior models are models in which the primary explanatory or predictive object is the behavior of an individual agent—its policy, trajectory, internal drivers, or deliberative process—rather than only aggregate system outputs. In recent work, the term denotes a surrogate policy learned from observed states and actions and exposed through a compact, locally interpretable behavior representation (Zhang et al., 2023); a framework that targets the conditional next-state distribution of each agent directly in agent-based models (Cozzi et al., 27 May 2025); a behavior model executed for each individual traffic participant in closed-loop simulation (Konstantinidis et al., 5 Dec 2025); and a framework that treats artificial agents as latent generative hypotheses about cognitive mechanisms (Ostwald et al., 30 Apr 2026). Across these uses, the common theme is that behavior is modeled at the level at which an agent observes, decides, acts, and adapts.

1. Formal scope and core objects

A standard formulation places the agent in sequential decision-making. At time tt, the agent observes a state sts_t, selects an action ata_t, and transitions to st+1s_{t+1}, with actions sampled from π(at∣st)\pi(a_t \mid s_t) and trajectories written as τ=s0,a0,s1,a1,…,st,at\tau = s_0, a_0, s_1, a_1, \dots, s_t, a_t. In this setting, an agent-level behavior model is not limited to a single input-output explanation; it can instead model the policy around a state or the decision logic that plausibly produced the action (Zhang et al., 2023).

For LLM-based agents, a trajectory may be represented as an ordered interaction history Ht=(o0,a0,o1,a1,…,ot)H_t = (o_0, a_0, o_1, a_1, \ldots, o_t), or as a sequence of temporally ordered components C=(C1,C2,…,C2T)C = (C_1, C_2, \ldots, C_{2T}), where each component is typed as {USER,THOUGHT,TOOL,OBS,MEMORY}\{\text{USER}, \text{THOUGHT}, \text{TOOL}, \text{OBS}, \text{MEMORY}\}. This representation makes explicit that agent behavior depends not only on current context, but on the order of tool calls, memory updates, and intermediate thoughts (Qian et al., 21 Jan 2026).

Controlled behavioral evaluation introduces a different but compatible formal object. ABxLab defines an environment E=⟨S,A,O,T,I⟩\mathcal{E} = \langle \mathcal{S}, \mathcal{A}, \mathcal{O}, \mathcal{T}, \mathcal{I} \rangle, with intervention functions sts_t0 that alter what the agent sees before it is observed. The operational move is to intervene on the observation stream itself, so that the model of behavior is defined through systematic changes in option attributes and persuasive cues rather than only through task completion (Cherep et al., 30 Sep 2025).

In cognitive behavioral science, the same topic is formalized as a joint probability model linking task variables, latent agent variables, observed actions, and rewards. For the Gabor task, the factorization is

sts_t1

Here the agent is a latent construct: it generates internal beliefs and decisions, but the experimenter sees only the actions (Ostwald et al., 30 Apr 2026).

2. Representational families

The literature instantiates agent-level behavior models through several distinct representational choices.

Representation Core object Representative use
Interpretable surrogate policy Decision tree and local decision path Behavior explanation (Zhang et al., 2023)
Adaptive stochastic process MCRE and higher-dimensional homogeneous Markov chain Agent behavior prediction (Tian et al., 2014)
Structural decision model Myopic utility maximization with latent motivational state Behavioral analytics (Mintz et al., 2017)
Learned nonstrategic choice Weighted level-0 distribution over actions Behavioral game theory (Wright et al., 2016)
Psychological or protocol-level model “Inner parliament” or finite-state behavior specification Human simulation and LLM agents (Hu et al., 4 Nov 2025, Crouse et al., 2023)

In explanation-oriented work, the representation is a decision-tree surrogate sts_t2 distilled from sampled rollouts, together with a state-specific decision path sts_t3. This path is the behavior representation: it is local, compact, and interpretable because each tree node corresponds to a human-readable condition (Zhang et al., 2023).

In platform settings with adaptive human or strategic agents, the behavior data are modeled as non-i.i.d. sequences generated by a Markov Chain in Random Environments. The state variables are joint behavior sts_t4, feedback sts_t5, and random user factors sts_t6, with transition sts_t7. The key technical step is to augment the state space to sts_t8, yielding a finite, time-homogeneous Markov chain that admits stationarity and uniform convergence analysis (Tian et al., 2014).

In intervention and incentive design, the model is often structural and agent-specific. The behavioral analytics framework assumes a discrete-time system with state sts_t9, motivational state ata_t0, action ata_t1, and incentive ata_t2, with dynamics ata_t3 and ata_t4. The myopia assumption gives

ata_t5

so the same individualized model supports estimation, prediction, and optimization (Mintz et al., 2017).

A classical antecedent in contract theory defines agent behavior through expected payoff

ata_t6

with derivative ata_t7 called the motivation function and second derivative ata_t8 called the persistence function. This broadens the standard assumption that effort is always a pure loss and makes risk contract-dependent rather than a fixed trait of the agent (Gutiérrez et al., 2011).

Behavioral game theory contributes another representational lineage. Level-0 behavior is modeled as a learnable distribution over actions based on features such as maxmax payoff, maxmin payoff, minmin unfairness, and max symmetric, rather than as a uniform distribution ata_t9. The recommended model is a linear weighting of features that requires the estimation of four weights (Wright et al., 2016).

Psychological and protocol-level models replace utility or state-space dynamics with explicit internal roles or explicit behavioral grammars. The psychological simulation system models a person as an “inner parliament” of agents such as Threat-Avoidance, Math-Anxiety, Goal-Pursuit, Procedural-Fluency, and Self-Efficacy (Hu et al., 4 Nov 2025). The formal specification framework models an LLM agent as a finite-state machine st+1s_{t+1}0 together with a declarative behavior formula expressed with operators such as next, until, and or (Crouse et al., 2023).

3. Learning and inference from observations, logs, and video

A major line of work learns behavior models from observed trajectories only. In the behavior-explanation setting, the policy is distilled into a decision tree using imitation learning, specifically DAgger: st+1s_{t+1}1 The explanation system is therefore independent of the agent’s internal representation and can be used for any policy from which trajectories can be sampled (Zhang et al., 2023).

Agent-based-model surrogates adopt a micro-level objective rather than a system-level one. The Graph Diffusion Network learns the conditional next-state distribution of each agent directly,

st+1s_{t+1}2

using a message-passing graph neural network to encode local interactions and a conditional diffusion model to capture behavioral stochasticity. Training is performed on a “ramification” dataset built from a main branch and many sibling branches, with denoising loss

st+1s_{t+1}3

The reported experiments show that the surrogate replicates individual-level patterns and forecasts emergent dynamics in Schelling’s segregation model and a Predator-Prey ecosystem (Cozzi et al., 27 May 2025).

Video-based behavior learning replaces simulator rollouts with persistent reconstruction. Agent-to-Sim first builds a persistent spacetime 4D representation from monocular RGBD smartphone videos, with a canonical structure st+1s_{t+1}4 and time-varying structure st+1s_{t+1}5. It then trains a hierarchical diffusion model over goal, path, and full body motion, conditioned on egocentric scene, observer, and past-motion codes. On the cat split, the main behavior metrics with st+1s_{t+1}6 samples are Goal st+1s_{t+1}7 m, Path st+1s_{t+1}8 m, Orientation st+1s_{t+1}9 rad, and Joint angles π(at∣st)\pi(a_t \mid s_t)0 rad (Yang et al., 2024).

Policy ABMs use RL as a behavior model when agents are treated as adaptive, reward-seeking decision-makers. The interface is gym-style,

π(at∣st)\pi(a_t \mid s_t)1

and the paper studies REINFORCE, REINFORCE with baseline, Actor-Critic, and a multi-agent actor-critic adaptation in the minority game and an influenza-transmission ABM. In the flu case, after about 750 training iterations with actor-critic, the RL agent improves its outcomes and on average outperforms the default behavioral model by a few percentage points over 100 seasons (Osoba et al., 2020).

A different simulation-based inference scheme is agent-centric Monte Carlo cognition. Here the agent’s “thinking” is another ABM: at each tick, the agent initializes a cognitive model from local observations, runs a settable number of rollouts for each candidate action, computes discounted reward based on change in energy, and executes the action with the highest mean reward. In the Wolf-Sheep Predation example, increasing the number of rollouts improves performance monotonically, and rollout length is most effective up to about three ticks (Head et al., 2018).

4. Explanation, attribution, and procedural analysis

One prominent use of agent-level behavior models is natural-language explanation grounded in a surrogate policy. The explanation pipeline is: π(at∣st)\pi(a_t \mid s_t)2 The prompt has four parts: a description of the environment and action semantics, a description of what the behavior representation means, in-context examples, and the specific behavior representation and action to explain. In an IRB-approved study with 32 participants, explanations generated from the behavior representation were significantly preferred over template-based explanations (π(at∣st)\pi(a_t \mid s_t)3, one-tailed binomial test), and participants did not prefer human explanations over the proposed method. In a hallucination analysis over 30 generated explanations, the behavior-representation method produced significantly fewer hallucinations than both baselines (Zhang et al., 2023).

General agentic attribution shifts the question from failure localization to action explanation regardless of task outcome. At component level, it computes

π(at∣st)\pi(a_t \mid s_t)4

so a large positive temporal gain marks a decision driver. At sentence level, it scores sentences with

π(at∣st)\pi(a_t \mid s_t)5

where probability drop measures necessity and probability hold measures sufficiency. Using human-annotated ground truth sentences, the default Prob. Drop&Hold method achieves Hit@1 π(at∣st)\pi(a_t \mid s_t)6, Hit@3 π(at∣st)\pi(a_t \mid s_t)7, and Hit@5 π(at∣st)\pi(a_t \mid s_t)8, outperforming LOO, ContextCite, and Saliency within the same hierarchical framework (Qian et al., 21 Jan 2026).

Coding-agent analysis treats an agent run as a trajectory through a code state-space. The Simple Strands Agent paper formalizes the “intent-execution gap” between what the model intends and what the harness executes, defines repository states π(at∣st)\pi(a_t \mid s_t)9, and measures live progress with a recall-style divergence

τ=s0,a0,s1,a1,…,st,at\tau = s_0, a_0, s_1, a_1, \dots, s_t, a_t0

Across SWE-Bench-Verified, SWE-Bench-Pro, and Terminal-Bench-2, the study analyzes about 138k trajectories. Its core empirical claim is that similar pass@1 scores can conceal very different edit frequency, testing activity, phase composition, and backtracking behavior (Gupta et al., 16 Jun 2026).

A complementary line represents trajectories themselves as procedures. “Trajectories as programs” induces an emergent vocabulary over action sequences with BPE-style merging, selects vocabulary size by V-measure, and reports a stable choice at τ=s0,a0,s1,a1,…,st,at\tau = s_0, a_0, s_1, a_1, \dots, s_t, a_t1 with V-measure τ=s0,a0,s1,a1,…,st,at\tau = s_0, a_0, s_1, a_1, \dots, s_t, a_t2. A probe over procedural fingerprints attributes unseen trajectories to the correct agent at 85.7% accuracy, against an 11.1% chance baseline for ten agents. ProcGrep, the structural querying library introduced in the same work, attains mean F1 τ=s0,a0,s1,a1,…,st,at\tau = s_0, a_0, s_1, a_1, \dots, s_t, a_t3 with latency τ=s0,a0,s1,a1,…,st,at\tau = s_0, a_0, s_1, a_1, \dots, s_t, a_t4 on its benchmarked structural queries, whereas Claude Sonnet 4.6 obtains F1 τ=s0,a0,s1,a1,…,st,at\tau = s_0, a_0, s_1, a_1, \dots, s_t, a_t5 with latency τ=s0,a0,s1,a1,…,st,at\tau = s_0, a_0, s_1, a_1, \dots, s_t, a_t6 s and GPT-4o obtains F1 τ=s0,a0,s1,a1,…,st,at\tau = s_0, a_0, s_1, a_1, \dots, s_t, a_t7 with latency τ=s0,a0,s1,a1,…,st,at\tau = s_0, a_0, s_1, a_1, \dots, s_t, a_t8 s (Oderinwale, 15 Jun 2026).

This suggests that explanation is no longer restricted to post hoc verbal rationalization. In this literature it includes local surrogate policies, historical attribution over tool and memory traces, and procedural fingerprints over full trajectories.

5. Simulation, intervention, and human-centered applications

In multi-agent decision support, agent-level behavior models can drive personalized interventions. The behavioral analytics framework organizes the workflow into three steps: develop a behavioral model, estimate behavioral model parameters and predict future decisions, and optimize a set of costly incentives. The single-agent 2SSA algorithm first computes a MAP estimate of τ=s0,a0,s1,a1,…,st,at\tau = s_0, a_0, s_1, a_1, \dots, s_t, a_t9 and then optimizes future incentives; the multi-agent ABMA algorithm decomposes the global problem into one MAP estimation per agent, one MILP per agent per discrete incentive type, and one master ILP. In a clinically supervised mobile weight-loss intervention, the framework maintains efficacy comparable to the original trial while reducing the required clinical-visit resources by up to about 60% (Mintz et al., 2017).

Driving simulation treats the behavior model as the policy executed for each traffic participant at every simulation step. The instance-centric design represents each traffic participant and each map element in its own local coordinate frame, uses a query-centric symmetric context encoder with relative positional encodings, and learns behavior with Adversarial Inverse Reinforcement Learning plus an adaptive reward offset Ht=(o0,a0,o1,a1,…,ot)H_t = (o_0, a_0, o_1, a_1, \ldots, o_t)0. The paper reports that the small model with only 59K parameters still outperforms all baselines on most metrics, that training time is reduced by up to 66.4% compared to agent-centric baselines, and that peak inference throughput reaches about 88k and 358k ISPS, up to 13.2× improvement over agent-centric baselines (Konstantinidis et al., 5 Dec 2025).

Human behavior simulation uses a more explicit internal architecture. The multi-agent psychological simulation system models a person as an “inner parliament” of psychological agents corresponding to constructs such as Threat-Avoidance, Math-Anxiety, Spatial-Reasoning, Goal-Pursuit, Procedural-Fluency, and Self-Efficacy. Behavior generation proceeds through initial positions, debate and adjustment over several rounds, and final synthesis. The system’s “Peek Into the Brain” view exposes the visible transcript of internal deliberation, and the paper frames the goal not as task optimality but as psychological authenticity in teacher training, psychological research, and related simulations (Hu et al., 4 Nov 2025).

Policy-focused ABMs offer another application domain. In the minority game and influenza-transmission case studies, RL behavioral models are used as adaptive, reward-seeking or reward-maximizing agents. The experiments show that such agents can learn to outperform the default adaptive behavioral models in the two ABMs examined, while synchronization analysis in the flu setting finds no strong synchronization, with correlations capped at about 10% magnitude (Osoba et al., 2020).

6. Evaluation, governance, and recurrent tensions

Behavioral evaluation frameworks argue that competence alone is insufficient. ABxLab turns websites into behavioral experiments by manipulating prices, ratings, order, and synthetic nudges in a 2AFC shopping environment. The benchmark runs over 80,000 experiments across 17 models, over about 2.5B tokens and about 400k requests. In the Original condition, higher ratings increased selection probability for 14 of 17 models with effects typically 30–80 pp, and the strongest effect was 81.2 pp for o4-mini. Price effects in the Original condition were about 15–34 pp for 13 of 17 models, while nudges shifted choices by about 10–60 pp on average when price and ratings were matched. Humans were far less sensitive overall, with order about 4 pp, rating about 5 pp, price about 9.4 pp, and nudge about 9.9 pp (Cherep et al., 30 Sep 2025).

Governance-oriented work extends the unit of analysis from local choice to the full lifecycle of network action. The “Network Behavior Lifecycle” decomposes behavior into Target Confirmation, Information Gathering, Reasoning Process, Decision Mechanism, Action Execution, and Feedback Acquisition. The proposed A4A paradigm uses agents to regulate, monitor, or evaluate other agents, while HABD compares humans and agents along five dimensions: decision mechanism, execution efficiency, intention-behavior consistency, behavioral inertia, and irrational patterns. In the red-team case, PentAGI required 2,000,000 GPT-4o tokens in zero-shot mode and 500,000 tokens with structured CoT prompts; in the blue-team case, the defensive coding comparison reports total time 68.2 s for the agent and 615 s for the human (Zhang et al., 20 Aug 2025).

Behavior can also be engineered declaratively rather than inferred post hoc. The formal specification framework lets a user define states and prompts, write a behavior formula such as Ht=(o0,a0,o1,a1,…,ot)H_t = (o_0, a_0, o_1, a_1, \ldots, o_t)1 for ReACT-like behavior, and compile the specification into a decoding monitor that validates, truncates, and repairs generated text through valid state prefixing. The same framework reproduces ReACT, ReWOO, Reflexion, direct answering, and chain-of-thought, and introduces PASS, a Plan-Act-Summarize-Solve agent. On HotpotQA, TriviaQA, and GSM8K, PASS is best among agent-based methods on HotpotQA and TriviaQA, while ReACT is best among agent-based methods on GSM8K (Crouse et al., 2023).

Several papers also make the limits of the field explicit. General agentic attribution argues that prior work predominantly focuses on failure attribution and is therefore blind to a correct but poorly justified answer, memory-induced bias, tool-conditioned hallucination, or overgeneralization from a prior success case (Qian et al., 21 Jan 2026). The coding-agent trajectory work argues that pass@1 is necessary but insufficient because similar solve rates can conceal large differences in trajectory quality and because public containers can leak future git history (Gupta et al., 16 Jun 2026). The psychological simulation paper is conceptual and architectural and does not provide a detailed mathematical formalism, explicit equations, or pseudo-code for the deliberation algorithm, while the governance paper describes HABD as qualitative and preliminary rather than fully formalized (Hu et al., 4 Nov 2025, Zhang et al., 20 Aug 2025).

Taken together, these formulations indicate that agent-level behavior models are not a single model class. They are a research program centered on the behavior of the individual agent as the unit of explanation, prediction, intervention, or governance. This suggests a broad convergence: surrogate policies, stochastic process models, utility-based decision systems, procedural trajectory models, and psychologically grounded simulators are all being used to ask the same technical question—how an agent behaves, why it behaves that way, and how that behavior should be evaluated or controlled.

Definition Search Book Streamline Icon: https://streamlinehq.com
References (19)

Topic to Video (Beta)

No one has generated a video about this topic yet.

Whiteboard

No one has generated a whiteboard explanation for this topic yet.

Follow Topic

Get notified by email when new papers are published related to Agent-Level Behavior Models.