Isbn: 9786139274529 - temporal difference learning: reinforcement learning, monte carlo method, dynamic programming, bootstrap aggregating, bellman equation (2 resultados)

- Tapa blanda
- Impresión bajo demanda
Librería: AHA-BUCH GmbH, Einbeck, AlemaniaAHA-BUCH GmbH
Contactar con el vendedorVendedor de 5 estrellasCondición: Nuevo
EUR 137,63
Envío por EUR 35,00Se envía de Alemania a Estados Unidos de AmericaCantidad disponible: 1 disponible
Taschenbuch. Condición: Neu. nach der Bestellung gedruckt Neuware - Printed after ordering - Please note that the content of this book primarily consists of articles available from Wikipedia or other free sources online. Temporal difference (TD) learning is a prediction method. It has been mostly used for solving the reinforcement learning problem. 'TD learning is a combination of Monte Carlo ideas and dynamic programming (DP) ideas.' TD resembles a Monte Carlo method because it learns by sampling the environment according to some policy. TD is related to dynamic programming techniques because it approximates its current estimate based on previously learned estimates (a process known as bootstrapping). The TD learning algorithm is related to the temporal difference model of animal learning.…

- Tapa blanda
- Impresión bajo demanda
Librería: preigu, Osnabrück, Alemaniapreigu
Contactar con el vendedorVendedor de 5 estrellasCondición: Nuevo
EUR 109,85
Envío por EUR 70,00Se envía de Alemania a Estados Unidos de AmericaCantidad disponible: 5 disponibles
Taschenbuch. Condición: Neu. Temporal Difference Learning | Reinforcement learning, Monte Carlo method, Dynamic programming, Bootstrap aggregating, Bellman equation | Klaas Apostol | Taschenbuch | Englisch | 2026 | OmniScriptum | EAN 9786139274529 | Verantwortliche Person für die EU: preigu GmbH & Co. KG, Lengericher Landstr. 19, 49078 Osnabrück, mail[at]preigu[dot]de | Anbieter: preigu Print on Demand. …