Abstract
The standard approach when using a Markov decision process to find an optimal policy is to assume a fixed profit or cost structure. However, in many applied problems it may be not possible to determine profits or costs associated with all states or actions. In such cases we propose the use of taboo first passage reward and taboo first passage time as objectives. In this paper we investigate problems related to optimizing aforementioned taboo measures and we provide two examples.