Peshkin leonid (8 Ergebnisse)

- Softcover
Anbieter: Rarewaves.com USA, London, LONDO, Vereinigtes KönigreichRarewaves.com USA
Verkäufer/-in kontaktierenVerkäufer/-in mit 5 SternenZustand: Neu
EUR 61,85
Versand gratisVersand von Vereinigtes Königreich nach USAAnzahl: Mehr als 20 verfügbar
Paperback. Zustand: New.

- Softcover
Anbieter: Ria Christie Collections, Uxbridge, Vereinigtes KönigreichRia Christie Collections
Verkäufer/-in kontaktierenVerkäufer/-in mit 5 SternenZustand: Neu
EUR 62,35
EUR 10,92 VersandVersand von Vereinigtes Königreich nach USAAnzahl: Mehr als 20 verfügbar
Zustand: New. In English.

- Softcover
Anbieter: moluna, Greven, Deutschlandmoluna
Verkäufer/-in kontaktierenVerkäufer/-in mit 5 SternenZustand: Neu
EUR 62,09
EUR 48,99 VersandVersand von Deutschland nach USAAnzahl: Mehr als 20 verfügbar
Zustand: New. Today we live in the world which is very much aman-made or artificial. In such a world there aremany systems and environments, both real andvirtual, which can be very well described by formalmodels. This creates an opportunity for de.

- Softcover
Anbieter: preigu, Osnabrück, Deutschlandpreigu
Verkäufer/-in kontaktierenVerkäufer/-in mit 5 SternenZustand: Neu
EUR 51,10
EUR 70,00 VersandVersand von Deutschland nach USAAnzahl: 5 verfügbar
Taschenbuch. Zustand: Neu. Reinforcement Learning from Scarce Experience via Policy Search | Learning to Act by Reasoning about Trial and Error in Uncertain Environment | Leonid Peshkin | Taschenbuch | Kartoniert / Broschiert | Englisch | 2013 | VDM Verlag Dr. Müller | EAN 9783639088038 | Verantwortliche Person für die EU: OmniScriptum GmbH & Co. KG, Bahnhofstr. 28, 66111 Saarbrücken, info[at]akademikerverlag[dot]de | Anbieter: preigu.…

- Softcover
Anbieter: Rarewaves.com UK, London, Vereinigtes KönigreichRarewaves.com UK
Verkäufer/-in kontaktierenVerkäufer/-in mit 5 SternenZustand: Neu
EUR 59,45
EUR 75,77 VersandVersand von Vereinigtes Königreich nach USAAnzahl: Mehr als 20 verfügbar
Paperback. Zustand: New.

- Softcover
- Print-on-Demand
Anbieter: PBShop.store US, Wood Dale, IL, USAPBShop.store US
Verkäufer/-in kontaktierenVerkäufer/-in mit 5 SternenZustand: Neu
EUR 62,31
Versand gratisVersand innerhalb von USAAnzahl: Mehr als 20 verfügbar
PAP. Zustand: New. New Book. Shipped from UK. THIS BOOK IS PRINTED ON DEMAND. Established seller since 2000.

- Softcover
- Print-on-Demand
Anbieter: PBShop.store UK, Fairford, GLOS, Vereinigtes KönigreichPBShop.store UK
Verkäufer/-in kontaktierenVerkäufer/-in mit 5 SternenZustand: Neu
EUR 59,00
EUR 4,85 VersandVersand von Vereinigtes Königreich nach USAAnzahl: Mehr als 20 verfügbar
PAP. Zustand: New. New Book. Delivered from our UK warehouse in 4 to 14 business days. THIS BOOK IS PRINTED ON DEMAND. Established seller since 2000.

- Softcover
- Print-on-Demand
Anbieter: AHA-BUCH GmbH, Einbeck, DeutschlandAHA-BUCH GmbH
Verkäufer/-in kontaktierenVerkäufer/-in mit 5 SternenZustand: Neu
EUR 62,05
EUR 35,00 VersandVersand von Deutschland nach USAAnzahl: 2 verfügbar
Taschenbuch. Zustand: Neu. nach der Bestellung gedruckt Neuware - Printed after ordering - Today we live in the world which is very much aman-made or artificial. In such a world there aremany systems and environments, both real andvirtual, which can be very well described by formalmodels. This creates an opportunity for developing a'synthetic intelligence' - artificial systemswhich cohabit these environments with human beings and carry out some useful function. In this book we address some aspects of thisdevelopment in the framework of reinforcementlearning, learning how to map sensations to actions,by trial and error from feedback. In some challengingcases, actions may affect not only the immediatereward, but also the next sensationand all subsequent rewards. The general task ofreinforcement learning stated in a traditional way isunreasonably ambitious for these two characteristics:search by trial-and-error and delayed reward. Weinvestigate general ways of breaking the task ofdesigning a controller down to more feasiblesub-tasks which are solved independently. We proposeto consider both taking advantage of past experienceby reusing parts of other systems, and facilitatingthe learning phase by employing a bias in initialconfiguration.…