Simulation-based Algorithms for Markov Decision Processes
☆☆☆☆☆
(0 reviews)
Lowest price (incl. delivery)
12 011,00 JPY
Typical price2 075,66 PLN
Lowest (90 days)75,50 PLN
Offers3
Last updated2 天前
Price history (90 days)
Full history
2026-08-08
2026-08-15
| 更新时间 | 价格 |
|---|---|
| 2026-08-08 | 88,80 |
| 2026-08-15 | 75,50 |
| 卖家 | Product price | Delivery | 总计 | 可用性 | Updated | |
|---|---|---|---|---|---|---|
| SP SpringerNatureLink Shop INT | 12 011,00 JPY | free | 12 011,00 JPY | 可购买 | 2 天前 | View offer |
| SP SpringerNatureLink Shop INT | 88,80 EUR | 19,00 EUR | 107,80 EUR | 可购买 | 2 天前 | View offer |
| SP Springer Nature Author | 112,00 EUR | free | 112,00 EUR | 可购买 | 1 周前 | View offer |
价格和库存可能会有变动。 最后更新: 15.08.2026 07:59.
EAN
9781846286902
Springer Nature
0,0
☆☆☆☆☆
0 reviews
5★
0%
4★
0%
3★
0%
2★
0%
1★
0%
Product reviews
No reviews yet — be the first!
Often, real-world problems modeled by Markov decision processes (MDPs) are difficult to solve in practise because of the curse of dimensionality. In others, explicit specification of the MDP model parameters is not feasible, but simulation samples are available. For these settings, various sampling and population-based numerical algorithms for computing an optimal solution in terms of a policy and/or value function have been developed recently. Here, this state-of-the-art research is brought together in a way that makes it accessible to researchers of varying interests and backgrounds. Many specific algorithms, illustrative numerical examples and rigorous theoretical convergence results are provided. The algorithms differ from the successful computational methods for solving MDPs based on neuro-dynamic programming or reinforcement learning. The algorithms can be combined with approximate dynamic programming methods that reduce the size of the state space and ameliorate the effects of dimensionality.