cs.NESep 28, 2026

Massively Parallel Reinforcement Learning with a Chaotic Reconfigurable Clockless Chip

Authors: Eric Oliveira-Gomes, Damien Rontani

Organizations: CentraleSup´elec and Universit´e de Lorraine, LMOPS UR4423 Laboratory, Metz F-57070, France

Abstract

Hardware accelerators based on physical dynamical systems offer an attractive route toward energy-efficient reinforcement learning applications. However, their scalability is challenging because it requires many statistically independent entropy sources. Here, we introduce a quasi-analog decision-making architecture based on asynchronous Boolean networks (or lattices) implemented on a clockless reconfigurable chip. Each node in the network consists of a single logic element that acts as an autonomous entropy source. This architecture gives rise to distributed Boolean chaos, in which a spatially coupled network generates parallel streams of chaotic Boolean transitions with very low statistical dependence between nodes. We experimentally demonstrate parallel decision-making on a 1024-armed bandit problem, which is beyond the scale of previous hardware implementations, while significantly improving power-law scaling performance. Separately, we scale the proposed entropy source to 5120 parallel channels, yielding an aggregate sample generation rate of 2.14 TS/s. Our solution is implemented on a commercial reconfigurable CMOS chip and offers high integration density and ease of programmability. Our results pave the way for using distributed Boolean chaos as a valuable hardware substrate for large-scale reinforcement learning and for the development of fully integrated, high-throughput decision-making accelerators.

Figures & tables

Explore similar work

CardsList
  1. Scalable neuromorphic computing from autonomous spiking dynamics in a clockless reconfigurable chip

    May 15, 2026Eric Oliveira Gomes, Damien RontaniNeuromorphic ComputingTime-To-First-Spike

  2. Neuromorphic Pseudo-Random Number Generators with a Low Power Hardware Implementation

    Sep 30, 2026Jafar Shamsi, Navid Akbari, Sonia Sennik +2Neuromorphic ComputingField-Programmable Gate Arrays

  3. On Distributional Reinforcement Learning in Chaotic Dynamical Systems

    May 28, 2026James Rudd-Jones, Mirco Musolesi, María Pérez-OrtizDistributional Reinforcement LearningChaos