A survey of multi-objective sequential decision-making

doi:10.1613/JAIR.3987

Open AccessJournal ArticleDOI

A survey of multi-objective sequential decision-making

Diederik M. Roijers, +3 more

- 01 Oct 2013 -

Journal of Artificial Intelligence Resea...

- Vol. 48, Iss: 1, pp 67-113

TLDR

This article surveys algorithms designed for sequential decision-making problems with multiple objectives and proposes a taxonomy that classifies multi-objective methods according to the applicable scenario, the nature of the scalarization function, and the type of policies considered.

Abstract:

Sequential decision-making problems with multiple objectives arise naturally in practice and pose unique challenges for research in decision-theoretic planning and learning, which has largely focused on single-objective settings. This article surveys algorithms designed for sequential decision-making problems with multiple objectives. Though there is a growing body of literature on this subject, little of it makes explicit under what circumstances special methods are needed to solve multi-objective problems. Therefore, we identify three distinct scenarios in which converting such a problem to a single-objective one is impossible, infeasible, or undesirable. Furthermore, we propose a taxonomy that classifies multi-objective methods according to the applicable scenario, the nature of the scalarization function (which projects multi-objective values to scalar ones), and the type of policies considered. We show how these factors determine the nature of an optimal solution, which can be a single policy, a convex hull, or a Pareto front. Using this taxonomy, we survey the literature on multi-objective methods for planning and learning. Finally, we discuss key applications of such methods and outline opportunities for future work.

A survey of multi-objective sequential decision-making

Citations

Analysing the Effects of Reward Shaping in Multi-Objective Stochastic Games

Reinforcement Learning from Simulated Environments: An Encoder Decoder Framework

Multi-Objective Simultaneous Optimistic Optimization

Reinforcement learning for Dialogue Systems optimization with user adaptation.

Actor-critic multi-objective reinforcement learning for non-linear utility functions

References

Dynamic Programming

Markov Decision Processes: Discrete Stochastic Dynamic Programming

Introduction to Reinforcement Learning

Evolutionary algorithms for solving multi-objective problems

Policy Gradient Methods for Reinforcement Learning with Function Approximation

Related Papers (5)

Reinforcement Learning: An Introduction

Human-level control through deep reinforcement learning

Markov Decision Processes: Discrete Stochastic Dynamic Programming

Multiobjective Reinforcement Learning: A Comprehensive Overview

Mastering the game of Go with deep neural networks and tree search