stat.MLOct 30, 2025

Action-Driven Processes for Continuous-Time Control

Authors: Ruimin He, Shaowei Lin

Organizations: Ukusan Pte Ltd, Singapore

Abstract

At the heart of reinforcement learning are actions -- decisions made in response to observations of the environment. Actions are equally fundamental in the modeling of stochastic processes, as they trigger discontinuous state transitions and enable the flow of information through large, complex systems. In this paper, we unify the perspectives of stochastic processes and reinforcement learning through action-driven processes, and illustrate their application to spiking neural networks. Leveraging ideas from control-as-inference, we show that minimizing the Kullback-Leibler divergence between a policy-driven true distribution and a reward-driven model distribution for a suitably defined action-driven process is equivalent to maximum entropy reinforcement learning.

Figures & tables

Explore similar work

CardsList
  1. From Ticks to Flows: Dynamics of Neural Reinforcement Learning in Continuous Environments

    Jun 2, 2026Saket Tiwari, Tejas Kotwal, George KonidarisSoft Actor-CriticOffline Reinforcement Learning

  2. Discretizing Reward Models

    Jun 19, 2026Vijay Viswanathan, Shiqi Wang, Devamanyu Hazarika +4Responses

  3. Diffusion-Augmented Markov Decision Processes for Maximum Entropy Reinforcement Learning

    Dec 1, 2025Sebastian Sanokowski, Kaustubh Patil, Majid KhadivDiffusion-Based Reinforcement Learning MethodsMarkov Decision Processes