cs.LGOct 8, 2026

When to Intervene? State-Aware Sparse Manipulation in Federated Reinforcement Learning

Authors: Shutong Zheng, Sijia Chen

Organizations: Sun Yat-sen University · The Hong Kong University of Science and Technology (Guangzhou)

Abstract

Federated reinforcement learning (FRL) enables distributed agents to collaboratively train decision-making policies, but its decentralized training process also exposes global policy learning to Byzantine manipulation. Existing poisoning attacks primarily focus on how to construct malicious updates, while trajectory-level intervention timing remains largely implicit. In sequential decision making, however, where an intervention is applied can alter subsequent trajectories and learning signals. Through controlled experiments, we find that changing the selected trajectory states materially alters attack efficacy even when the malicious-update construction is fixed. We therefore identify when as a distinct attack dimension and introduce the Viability-constrained Behavioral Steering Attack (V-BSA), which uses local policy uncertainty to select sparse intervention states and applies envelope-constrained behavioral steering. Across discrete-action benchmarks, V-BSA achieves substantial degradation against robust aggregators and ensemble defenses with only a fraction of the interventions used by dense poisoning, while revealing task- and aggregation-dependent boundaries. Overall, our results highlight intervention timing as a distinct dimension of sequential robustness in FRL. The code is available at https://github.com/Yodeesy/V-BSA

Figures & tables

Appendix figures & tables10 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. Constraint-Aware Aggregation for Federated Reinforcement Learning in Microgrid Energy Coordination

    Jul 14, 2026Usman Haider, Karl MasonFederated Learning AggregationFederated RL

  2. Efficient Federated RLHF via Zeroth-Order Policy Optimization

    Apr 20, 2026Deyi Wang, Qining Zhang, Lei YingFederated RLCommunication-Efficient Distributed Training

  3. Fed-CausalDiff: Decoupled Synchronization for Federated Do-Simulation and Policy Evaluation

    Jun 21, 2026Pengfei Li, Mohammad KhalilPolicy EvaluationCausal Inference