HOW TO MAXIMIZE REWARD RATE ON TWO VARIABLE‐INTERVAL PARADIGMS

  • Houston A
  • McNamara J
106Citations
Citations of this article
32Readers
Mendeley users who have this article in their library.
Get full text

Abstract

Without assuming any constraints on behavior, we derive the policy that maximizes overall reward rate on two variable‐interval paradigms. The first paradigm is concurrent variable time‐variable time with changeover delay. It is shown that for nearly all parameter values, a switch to the schedule with the longer interval should be followed immediately by a switch back to the schedule with the shorter interval. The matching law does not hold at the optimum and does not uniquely specify the obtained reward rate. The second paradigm is discrete trial concurrent variable interval‐variable interval. For given schedule parameters, the optimal policy involves a cycle of a fixed number of choices of the schedule with the shorter interval followed by one choice of the schedule with the longer interval. Molecular maximization sometimes results in optimal behavior.

Cite

CITATION STYLE

APA

Houston, A. I., & McNamara, J. (1981). HOW TO MAXIMIZE REWARD RATE ON TWO VARIABLE‐INTERVAL PARADIGMS. Journal of the Experimental Analysis of Behavior, 35(3), 367–396. https://doi.org/10.1901/jeab.1981.35-367

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free