Search arXivSearch

arXiv · 2407.07310

Optimal Sensor and Actuator Selection for Factored Markov Decision Processes: Complexity, Approximability and Algorithms

Abstract

Factored Markov Decision Processes (fMDPs) are a class of Markov Decision Processes (MDPs) in which the states (and actions) can be factored into a set of state (and action) variables and can be encoded compactly using a factored representation. In this paper, we consider a setting where the state of the fMDP is not directly observable, and the agent relies on a set of potential sensors to gather information. We formulate the problem of selecting a set of sensors for fMDPs (under a limited budget) to maximize the infinite-horizon discounted return provided by the optimal policy. We show the fundamental result that it is NP-hard to approximate this problem to within a factor of $n^{1-c}$ for any $c > 1$, where $n$ is the number of state variables. Our inapproximability results for sensor selection also extend to a general class of Partially Observable MDPs (POMDPs). We also consider the dual problem of budgeted actuator selection (at design-time) to maximize the expected return under the optimal policy, for which we establish similar inapproximability results. Finally, we consider a simple greedy algorithm and empirically show that, despite the lack of formal theoretical guarantees, it performs effectively in practice, achieving on average over $70\%$ of the optimal solution value across a variety of real-world and randomly generated problem instances.

Explore related subjects

Keep this discovery

BibTeXRIS

Jayanth Bhargav, Mahsa Ghasemi, Shreyas Sundaram. 2026-09-02. Optimal Sensor and Actuator Selection for Factored Markov Decision Processes: Complexity, Approximability and Algorithms. https://arxiv.org/abs/2407.07310

Cite the original work for its findings. Save a collection to share your selection of sources.

Discover connections

Connections use source metadata and explicit phrase matches, not verified experimental comparisons.

KEEP EXPLORING

Related papers

Equality cases of the Stanley--Yan log-concave matroid inequality

The \emph{Stanley--Yan} (SY) \emph{inequality} gives the ultra-log-concavity for the numbers of bases of a matroid which have given sizes of intersections with $k$ fixed disjoint sets. The inequality was proved by Stanley (1981) for regular matroids, and by Yan (2023) in full generality. In the original paper, Stanley asked for equality conditions of the SY~inequality, and proved total equality conditions for regular matroids in the case $k=0$. In this paper, we completely resolve Stanley's problem. First, we obtain an explicit description of the equality cases of the SY inequality for $k=0$, extending Stanley's results to general matroids and removing the ``total equality'' assumption. Second, for $k\ge 1$, we prove that the equality cases of the SY inequality cannot be described in a sense that they are not in the polynomial hierarchy unless the polynomial hierarchy collapses to a finite level.

math.CO

A simple derivation of the Kalman filter

In this lecture note, we present a concise and self-contained derivation of the discrete-time Kalman filter equations that requires only a basic understanding of least squares estimation. The treatment is designed to minimize mathematical overhead while preserving both rigor and generality.

math.OC