acceptodds
Under review as a conference paper at ICLR 2027

Behaviorally Silent Now, Causally Active Later: Deferred State Effects in Recurrent Policies for Reinforcement Learning

Abstract

Observation perturbations in reinforcement learning (RL) agents with recurrent policies can have lasting effects on both the policy's internal state and its subsequent behavior. However, when such delayed effects occur, it remains unclear whether they arise from the immediate action change induced by the perturbation or from perturbation information retained in the policy's internal state. To address this, we introduce current action equivalence, which separates these pathways by comparing internal states that produce the same current action. For the one-shot intervention on feedforward policies, reproducing the current action reproduces the downstream effect. In recurrent policies, however, real observation perturbations can induce internal state changes that remain behaviorally silent at the current step yet become increasingly expressed through future actions. We identify these changes as deferred state effects and show through interventions that they causally influence future actions, closed-loop trajectories, and task reward. The effect is consistent across perturbations, recurrent architectures, environments, and independently trained policies. These results show that the effects of a perturbation can remain behaviorally silent at the present step while continuing to shape future decisions in recurrent policies.

open until 14 Dec 2026

est. 32% chance this paper gets accepted at ICLR 2027.

Reject 68%Accept 32%

What do you think this paper will get?

All positions stay anonymous.

Related papers

Loading the map…

Discussion (0)

Sign in to comment.