Out of stock

This item is currently unavailable

Model-based Reinforcement Learning: A Survey (Foundations and Trends® in Machine Learning)

Out of stock

Price data last checked 11 day(s) ago - will refresh soon

View at Amazon

One email. No newsletter. No nudges.

Gone for 111 days. Could come back at any time — we're watching for you.

Out of stock 111 days · last price £50 · longest previous gap was 1 days

NEW HERE?

Amazon shows you one price. We show you all of them.

Tosheroon watches Amazon prices so you don't have to. Every product on Amazon has a price history — we make it visible. Set the price you'd actually pay, and we'll email you the second it gets there. No app, no account, one email.

WHAT'S ON THIS PAGE

↓ Price chart
when this has been cheap or pricey
↓ Forecast
where the price is heading next
↓ Statistics
all-time high & low, recent range
↑ Price alert
name your number, we'll email you

Price History & Forecast

Grey patches = out of stock. Cheaper = lower on the chart. Hover for exact prices.

Last 315 days · 315 data points (no recent data)

Historical
Generating forecast…
£76.45 £47.59 £53.88 £60.18 £66.48 £72.78 £79.07 02 September 2025 19 November 2025 06 February 2026 25 April 2026 13 July 2026

Price Distribution

Price distribution over 315 days • 6 price levels

Days at Price
Current Price
152 days · current 89 days 11 days 2 days 13 days 48 days 0 38 76 114 152 £50 £53 £55 £61 £71 £76 Days at Price

Price Analysis

Most common price: £50 (152 days, 48.3%)

Price range: £50 - £76

Price levels: 6 different prices over 315 days

Description

Sequential decision making, commonly formalized as Markov Decision Process (MDP) optimization, is an important challenge in artificial intelligence. Two key approaches to this problem are reinforcement learning (RL) and planning. This monograph surveys an integration of both fields, better known as model-based reinforcement learning. Model-based RL has two main steps: dynamics model learning and planning-learning integration. In this comprehensive survey of the topic, the authors first cover dynamics model learning, including challenges such as dealing with stochasticity, uncertainty, partial observability, and temporal abstraction. They then present a systematic categorization of planning-learning integration, including aspects such as: where to start planning, what budgets to allocate to planning and real data collection, how to plan, and how to integrate planning in the learning and acting loop. In conclusion the authors discuss implicit model-based RL as an end-to-end alternative for model learning and planning, and cover the potential benefits of model-based RL. Along the way, the authors draw connections to several related RL fields, including hierarchical RL and transfer learning. This monograph contains a broad conceptual overview of the combination of planning and learning for Markov Decision Process optimization. It provides a clear and complete introduction to the topic for students and researchers alike.

Product Specifications

Format
paperback
Domain
Amazon UK
Release Date
04 January 2023
Listed Since
09 January 2023

Barcode

No barcode data available