cs.AIJul 18, 2026

Expected Free Energy as Belief-Dependent Utility for rho-POMDPs

Authors: Patrick CooperAlvaro Velasquez

Organizations: Department of Computer Science, University of Colorado Boulder, Boulder, CO, USA

Abstract

An agent acting under partial observability must decide when to gather information and which observations are worth their cost. Standard POMDPs value information only through its eventual effect on reward. The ρρ-POMDP framework instead rewards uncertainty reduction directly, through a belief-dependent utility ρρ, but in practice both the choice of ρρ and the weight placed on it are tuned by hand for every task. We show that active inference removes this tuning entirely. Minimizing Expected Free Energy (EFE) is exactly equivalent to solving a ρρ-POMDP whose utility is expected information gain, and the exploration weight is fixed at w=1w=1 because the variational bound expresses pragmatic and epistemic value in the same units (nats). We prove this equivalence for observe-then-commit POMDPs and extend it to factored observation POMDPs, a broader class that covers interleaved observe-act problems such as non-destructive testing and mobile sensing, where gathering information leaves the hidden state unchanged. Experiments support the theory. Across environments ranging from the classic Tiger problem to RockSample and a new Structural Inspection benchmark with over 65,000 states, the untuned weight matches or outperforms reward-only planning at the same horizon, avoids the over-exploration of bonuses tuned per task, and sits near the reward-maximizing knee of the success-reward Pareto frontier. The practical payoff is an exploration objective that works out of the box. In applications such as fault detection and medical screening, where every test has a price and every missed fault has a cost, EFE supplies a belief-dependent utility that is derived rather than tuned.

Explore similar work

CardsList
  1. Expected Free Energy-based Planning as Variational Inference

    Jun 9, 2026Wouter W. L. Nuijten, Thijs van de Laar, Bert de VriesVariational InferenceClassical Planning

  2. What Type of Inference is Active Inference?

    Jun 3, 2026Wouter W. L. Nuijten, Mykola Lukashchuk, Thijs van de Laar +1Classical PlanningFree Energy Principle