Peshkin leonid (8 resultados)

- Tapa blanda
Librería: Rarewaves.com USA, London, LONDO, Reino UnidoRarewaves.com USA
Contactar con el vendedorVendedor de 5 estrellasCondición: Nuevo
EUR 61,81
Gastos de envío gratisSe envía de Reino Unido a Estados Unidos de AmericaCantidad disponible: Más de 20 disponibles
Paperback. Condición: New.

- Tapa blanda
Librería: Ria Christie Collections, Uxbridge, Reino UnidoRia Christie Collections
Contactar con el vendedorVendedor de 5 estrellasCondición: Nuevo
EUR 62,19
Envío por EUR 10,89Se envía de Reino Unido a Estados Unidos de AmericaCantidad disponible: Más de 20 disponibles
Condición: New. In English.

- Tapa blanda
Librería: moluna, Greven, Alemaniamoluna
Contactar con el vendedorVendedor de 5 estrellasCondición: Nuevo
EUR 62,09
Envío por EUR 48,99Se envía de Alemania a Estados Unidos de AmericaCantidad disponible: Más de 20 disponibles
Condición: New. Today we live in the world which is very much aman-made or artificial. In such a world there aremany systems and environments, both real andvirtual, which can be very well described by formalmodels. This creates an opportunity for de.

- Tapa blanda
Librería: preigu, Osnabrück, Alemaniapreigu
Contactar con el vendedorVendedor de 5 estrellasCondición: Nuevo
EUR 51,10
Envío por EUR 70,00Se envía de Alemania a Estados Unidos de AmericaCantidad disponible: 5 disponibles
Taschenbuch. Condición: Neu. Reinforcement Learning from Scarce Experience via Policy Search | Learning to Act by Reasoning about Trial and Error in Uncertain Environment | Leonid Peshkin | Taschenbuch | Kartoniert / Broschiert | Englisch | 2013 | VDM Verlag Dr. Müller | EAN 9783639088038 | Verantwortliche Person für die EU: OmniScriptum GmbH & Co. KG, Bahnhofstr. 28, 66111 Saarbrücken, info[at]akademikerverlag[dot]de | Anbieter: preigu.…

- Tapa blanda
Librería: Rarewaves.com UK, London, Reino UnidoRarewaves.com UK
Contactar con el vendedorVendedor de 5 estrellasCondición: Nuevo
EUR 59,29
Envío por EUR 75,57Se envía de Reino Unido a Estados Unidos de AmericaCantidad disponible: Más de 20 disponibles
Paperback. Condición: New.

- Tapa blanda
- Impresión bajo demanda
Librería: PBShop.store US, Wood Dale, IL, Estados Unidos de AmericaPBShop.store US
Contactar con el vendedorVendedor de 5 estrellasCondición: Nuevo
EUR 62,27
Gastos de envío gratisSe envía dentro de Estados Unidos de AmericaCantidad disponible: Más de 20 disponibles
PAP. Condición: New. New Book. Shipped from UK. THIS BOOK IS PRINTED ON DEMAND. Established seller since 2000.

- Tapa blanda
- Impresión bajo demanda
Librería: PBShop.store UK, Fairford, GLOS, Reino UnidoPBShop.store UK
Contactar con el vendedorVendedor de 5 estrellasCondición: Nuevo
EUR 58,85
Envío por EUR 4,84Se envía de Reino Unido a Estados Unidos de AmericaCantidad disponible: Más de 20 disponibles
PAP. Condición: New. New Book. Delivered from our UK warehouse in 4 to 14 business days. THIS BOOK IS PRINTED ON DEMAND. Established seller since 2000.

- Tapa blanda
- Impresión bajo demanda
Librería: AHA-BUCH GmbH, Einbeck, AlemaniaAHA-BUCH GmbH
Contactar con el vendedorVendedor de 5 estrellasCondición: Nuevo
EUR 62,05
Envío por EUR 35,00Se envía de Alemania a Estados Unidos de AmericaCantidad disponible: 2 disponibles
Taschenbuch. Condición: Neu. nach der Bestellung gedruckt Neuware - Printed after ordering - Today we live in the world which is very much aman-made or artificial. In such a world there aremany systems and environments, both real andvirtual, which can be very well described by formalmodels. This creates an opportunity for developing a'synthetic intelligence' - artificial systemswhich cohabit these environments with human beings and carry out some useful function. In this book we address some aspects of thisdevelopment in the framework of reinforcementlearning, learning how to map sensations to actions,by trial and error from feedback. In some challengingcases, actions may affect not only the immediatereward, but also the next sensationand all subsequent rewards. The general task ofreinforcement learning stated in a traditional way isunreasonably ambitious for these two characteristics:search by trial-and-error and delayed reward. Weinvestigate general ways of breaking the task ofdesigning a controller down to more feasiblesub-tasks which are solved independently. We proposeto consider both taking advantage of past experienceby reusing parts of other systems, and facilitatingthe learning phase by employing a bias in initialconfiguration.…