Home > Publications . Search All . Browse All . Country . Browse PSC Pubs . PSC Report Series

PSC In The News

RSS Feed icon

Edin and Shaefer's book a call to action for Americans to deal with poverty

Weir says pain may underlie rise in suicide and substance-related deaths among white middle-aged Americans

Weitzman says China's one-child policy has had devastating effects on first-born daughters


MCubed opens for new round of seed funding, November 4-18

PSC News, fall 2015 now available

Barbara Anderson appointed chair of Census Scientific Advisory Committee

John Knodel honored by Thailand's Chulalongkorn University

Next Brown Bag

Monday, Dec 7 at noon, 6050 ISR-Thompson
Daniel Eisenberg, "Healthy Minds Network: Mental Health among College-Age Populations"

Batch mode reinforcement learning based on the synthesis of artificial trajectories

Publication Abstract

Fonteneau, R., Susan A. Murphy, L. Wehenkel, and D. Ernst. 2013. "Batch mode reinforcement learning based on the synthesis of artificial trajectories." Annals of Operations Research, 208(1): 383-416.

In this paper, we consider the batch mode reinforcement learning setting, where the central problem is to learn from a sample of trajectories a policy that satisfies or optimizes a performance criterion. We focus on the continuous state space case for which usual resolution schemes rely on function approximators either to represent the underlying control problem or to represent its value function. As an alternative to the use of function approximators, we rely on the synthesis of "artificial trajectories" from the given sample of trajectories, and show that this idea opens new avenues for designing and analyzing algorithms for batch mode reinforcement learning.

DOI:10.1007/s10479-012-1248-5 (Full Text)

PMCID: PMC3773886. (Pub Med Central)

Browse | Search : All Pubs | Next