Csaba szepesvari (41 risultati)

Autore: 
Perfeziona con la Ricerca avanzata

Perfeziona la tua ricerca

  • Libri (41)

a

Fascia di prezzo personalizzata (EUR)

a

  • Lingua: Inglese

    Editore: Springer International Publishing AG, Cham, 2010

    303100423X / 9783031004230

    • Brossura

    Da: Grand Eagle Retail, Bensenville, IL, U.S.A.Grand Eagle Retail

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 34,42

     Spedizione gratuita 
    Spedito in U.S.A.

    Quantità: 1 disponibile

    Paperback. Condizione: new. Paperback. Reinforcement learning is a learning paradigm concerned with learning to control a system so as to maximize a numerical performance measure that expresses a long-term objective. What distinguishes reinforcement learning from supervised learning is that only partial feedback is given to the learner about the learner's predictions. Further, the predictions may have long term effects through influencing the future state of the controlled system. Thus, time plays a special role. The goal in reinforcement learning is to develop efficient learning algorithms, as well as to understand the algorithms' merits and limitations. Reinforcement learning is of great interest because of the large number of practical applications that it can be used to address, ranging from problems in artificial intelligence to operations research or control engineering. In this book, we focus on those algorithms of reinforcement learning that build on the powerful theory of dynamic programming. We give a fairly comprehensive catalog of learning problems, describe the core ideas, note a large number of state of the art algorithms, followed by the discussion of their theoretical properties and limitations. Table of Contents: Markov Decision Processes / Value Prediction Problems / Control / For Further Exploration Reinforcement learning is a learning paradigm concerned with learning to control a system so as to maximize a numerical performance measure that expresses a long-term objective. What distinguishes reinforcement learning from supervised learning is that only partial feedback is given to the learner about the learner's predictions. Shipping may be from multiple locations in the US or from the UK, depending on stock availability.…

  • Lingua: Inglese

    Editore: Springer, 2010

    303100423X / 9783031004230

    • Brossura

    Da: California Books, Miami, FL, U.S.A.California Books

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 38,45

     Spedizione gratuita 
    Spedito in U.S.A.

    Quantità: Più di 20 disponibili

    Condizione: New.

  • Lingua: Inglese

    Editore: Springer International Publishing AG, CH, 2010

    303100423X / 9783031004230

    • Brossura
    • Prima edizione

    Da: Rarewaves.com USA, London, LONDO, Regno UnitoRarewaves.com USA

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 39,01

     Spedizione gratuita 
    Spedito da Regno Unito a U.S.A.

    Quantità: Più di 20 disponibili

    Paperback. Condizione: New. 1st. Reinforcement learning is a learning paradigm concerned with learning to control a system so as to maximize a numerical performance measure that expresses a long-term objective. What distinguishes reinforcement learning from supervised learning is that only partial feedback is given to the learner about the learner's predictions. Further, the predictions may have long term effects through influencing the future state of the controlled system. Thus, time plays a special role. The goal in reinforcement learning is to develop efficient learning algorithms, as well as to understand the algorithms' merits and limitations. Reinforcement learning is of great interest because of the large number of practical applications that it can be used to address, ranging from problems in artificial intelligence to operations research or control engineering. In this book, we focus on those algorithms of reinforcement learning that build on the powerful theory of dynamic programming. We give a fairly comprehensive catalog of learning problems, describe the core ideas, note a large number of state of the art algorithms, followed by the discussion of their theoretical properties and limitations. Table of Contents: Markov Decision Processes / Value Prediction Problems / Control / For Further Exploration.…

  • Lingua: Inglese

    Editore: Springer, 2010

    303100423X / 9783031004230

    • Brossura

    Da: Basi6 International, Irving, TX, U.S.A.Basi6 International

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 46,24

     Spedizione gratuita 
    Spedito in U.S.A.

    Quantità: 1 disponibile

    Condizione: Brand New. New. US edition. Expediting shipping for all USA and Europe orders excluding PO Box. Excellent Customer Service.

  • Lingua: Inglese

    Editore: Springer, 2010

    303100423X / 9783031004230

    • Brossura

    Da: Ria Christie Collections, Uxbridge, Regno UnitoRia Christie Collections

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 38,52

    EUR 11,03 spedizione 
    Spedito da Regno Unito a U.S.A.

    Quantità: Più di 20 disponibili

    Condizione: New. In English.

  • Lingua: Inglese

    Editore: Springer-Verlag Berlin and Heidelberg GmbH & Co. KG, Berlin, 2011

    3642244114 / 9783642244117

    • Brossura

    Da: Grand Eagle Retail, Bensenville, IL, U.S.A.Grand Eagle Retail

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 56,41

     Spedizione gratuita 
    Spedito in U.S.A.

    Quantità: 1 disponibile

    Paperback. Condizione: new. Paperback. This book constitutes the refereed proceedings of the 22nd International Conference on Algorithmic Learning Theory, ALT 2011, held in Espoo, Finland, in October 2011, co-located with the 14th International Conference on Discovery Science, DS 2011. The 28 revised full papers presented together with the abstracts of 5 invited talks were carefully reviewed and selected from numerous submissions. The papers are divided into topical sections of papers on inductive inference, regression, bandit problems, online learning, kernel and margin-based methods, intelligent agents and other learning models. This book constitutes the refereed proceedings of the 22nd International Conference on Algorithmic Learning Theory, ALT 2011, held in Espoo, Finland, in October 2011, co-located with the 14th International Conference on Discovery Science, DS 2011. Shipping may be from multiple locations in the US or from the UK, depending on stock availability.…

  • Lingua: Inglese

    Editore: Springer, 2010

    303100423X / 9783031004230

    • Brossura

    Da: Books Puddle, Woodside, NY, U.S.A.Books Puddle

    Venditore con 4 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 53,59

    EUR 3,55 spedizione 
    Spedito in U.S.A.

    Quantità: 4 disponibili

    Condizione: New.

  • Lingua: Inglese

    Editore: Morgan & Claypool, 2010

    1608454924 / 9781608454921

    • Brossura

    Da: ThriftBooks-Dallas, Dallas, TX, U.S.A.ThriftBooks-Dallas

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Usato - Molto buono

    EUR 59,29

     Spedizione gratuita 
    Spedito in U.S.A.

    Quantità: 1 disponibile

    Paperback. Condizione: Very Good. No Jacket. May have limited writing in cover pages. Pages are unmarked. ~ ThriftBooks: Read More, Spend Less.

  • Lingua: Inglese

    Editore: Cambridge University Press, 2020

    1108486827 / 9781108486828

    • Rilegato

    Da: California Books, Miami, FL, U.S.A.California Books

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 66,82

     Spedizione gratuita 
    Spedito in U.S.A.

    Quantità: Più di 20 disponibili

    Condizione: New.

  • Lingua: Inglese

    Editore: Springer-Verlag Berlin and Heidelberg GmbH and Co. KG, DE, 2011

    3642244114 / 9783642244117

    • Brossura

    Da: Rarewaves.com USA, London, LONDO, Regno UnitoRarewaves.com USA

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 70,18

     Spedizione gratuita 
    Spedito da Regno Unito a U.S.A.

    Quantità: Più di 20 disponibili

    Paperback. Condizione: New. 2011th. This book constitutes the refereed proceedings of the 22nd International Conference on Algorithmic Learning Theory, ALT 2011, held in Espoo, Finland, in October 2011, co-located with the 14th International Conference on Discovery Science, DS 2011. The 28 revised full papers presented together with the abstracts of 5 invited talks were carefully reviewed and selected from numerous submissions. The papers are divided into topical sections of papers on inductive inference, regression, bandit problems, online learning, kernel and margin-based methods, intelligent agents and other learning models.…

  • Lingua: Inglese

    Editore: Springer, 2010

    303100423X / 9783031004230

    • Brossura

    Da: AHA-BUCH GmbH, Einbeck, GermaniaAHA-BUCH GmbH

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 34,46

    EUR 35,00 spedizione 
    Spedito da Germania a U.S.A.

    Quantità: 1 disponibile

    Taschenbuch. Condizione: Neu. Druck auf Anfrage Neuware - Printed after ordering - Reinforcement learning is a learning paradigm concerned with learning to control a system so as to maximize a numerical performance measure that expresses a long-term objective. What distinguishes reinforcement learning from supervised learning is that only partial feedback is given to the learner about the learner's predictions. Further, the predictions may have long term effects through influencing the future state of the controlled system. Thus, time plays a special role. The goal in reinforcement learning is to develop efficient learning algorithms, as well as to understand the algorithms' merits and limitations. Reinforcement learning is of great interest because of the large number of practical applications that it can be used to address, ranging from problems in artificial intelligence to operations research or control engineering. In this book, we focus on those algorithms of reinforcement learning that build on the powerful theory of dynamic programming. We give a fairly comprehensive catalog of learning problems, describe the core ideas, note a large number of state of the art algorithms, followed by the discussion of their theoretical properties and limitations. Table of Contents: Markov Decision Processes / Value Prediction Problems / Control / For Further Exploration.…

  • Condizione: Nuovo

    EUR 58,14

    EUR 18,23 spedizione 
    Spedito da Regno Unito a U.S.A.

    Quantità: 10 disponibili

    Paperback. Condizione: New.

  • Lingua: Inglese

    Editore: Springer, 2011

    3642244114 / 9783642244117

    • Brossura

    Da: Majestic Books, Hounslow, Regno UnitoMajestic Books

    Venditore con 4 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 68,82

    EUR 7,65 spedizione 
    Spedito da Regno Unito a U.S.A.

    Quantità: 4 disponibili

    Condizione: New. pp. 468 Illus.

  • Lingua: Inglese

    Editore: Cambridge University Press, GB, 2020

    1108486827 / 9781108486828

    • Rilegato

    Da: Rarewaves.com USA, London, LONDO, Regno UnitoRarewaves.com USA

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 78,54

     Spedizione gratuita 
    Spedito da Regno Unito a U.S.A.

    Quantità: Più di 20 disponibili

    Hardback. Condizione: New. Decision-making in the face of uncertainty is a significant challenge in machine learning, and the multi-armed bandit model is a commonly used framework to address it. This comprehensive and rigorous introduction to the multi-armed bandit problem examines all the major settings, including stochastic, adversarial, and Bayesian frameworks. A focus on both mathematical intuition and carefully worked proofs makes this an excellent reference for established researchers and a helpful resource for graduate students in computer science, engineering, statistics, applied mathematics and economics. Linear bandits receive special attention as one of the most useful models in applications, while other chapters are dedicated to combinatorial bandits, ranking, non-stationary problems, Thompson sampling and pure exploration. The book ends with a peek into the world beyond bandits with an introduction to partial monitoring and learning in Markov decision processes.…

  • Lingua: Inglese

    Editore: Springer, 2011

    3642244114 / 9783642244117

    • Brossura

    Da: Ria Christie Collections, Uxbridge, Regno UnitoRia Christie Collections

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 67,81

    EUR 13,28 spedizione 
    Spedito da Regno Unito a U.S.A.

    Quantità: Più di 20 disponibili

    Condizione: New. In English.

  • Lingua: Inglese

    Editore: Cambridge University Press, 2020

    1108486827 / 9781108486828

    • Rilegato

    Da: Majestic Books, Hounslow, Regno UnitoMajestic Books

    Venditore con 4 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 77,18

    EUR 7,65 spedizione 
    Spedito da Regno Unito a U.S.A.

    Quantità: 1 disponibile

    Condizione: New.

  • Lingua: Inglese

    Editore: Springer International Publishing AG, Cham, 2010

    303100423X / 9783031004230

    • Brossura

    Da: AussieBookSeller, Truganina, VIC, AustraliaAussieBookSeller

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 51,74

    EUR 32,88 spedizione 
    Spedito da Australia a U.S.A.

    Quantità: 1 disponibile

    Paperback. Condizione: new. Paperback. Reinforcement learning is a learning paradigm concerned with learning to control a system so as to maximize a numerical performance measure that expresses a long-term objective. What distinguishes reinforcement learning from supervised learning is that only partial feedback is given to the learner about the learner's predictions. Further, the predictions may have long term effects through influencing the future state of the controlled system. Thus, time plays a special role. The goal in reinforcement learning is to develop efficient learning algorithms, as well as to understand the algorithms' merits and limitations. Reinforcement learning is of great interest because of the large number of practical applications that it can be used to address, ranging from problems in artificial intelligence to operations research or control engineering. In this book, we focus on those algorithms of reinforcement learning that build on the powerful theory of dynamic programming. We give a fairly comprehensive catalog of learning problems, describe the core ideas, note a large number of state of the art algorithms, followed by the discussion of their theoretical properties and limitations. Table of Contents: Markov Decision Processes / Value Prediction Problems / Control / For Further Exploration Reinforcement learning is a learning paradigm concerned with learning to control a system so as to maximize a numerical performance measure that expresses a long-term objective. What distinguishes reinforcement learning from supervised learning is that only partial feedback is given to the learner about the learner's predictions. Shipping may be from our Sydney, NSW warehouse or from our UK or US warehouse, depending on stock availability.…

  • Lingua: Inglese

    Editore: Cambridge University Press, 2020

    1108486827 / 9781108486828

    • Rilegato

    Da: Ria Christie Collections, Uxbridge, Regno UnitoRia Christie Collections

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 69,19

    EUR 17,58 spedizione 
    Spedito da Regno Unito a U.S.A.

    Quantità: Più di 20 disponibili

    Condizione: New. In English.

  • Lingua: Inglese

    Editore: Cambridge University Press CUP, 2020

    1108486827 / 9781108486828

    • Rilegato

    Da: Books Puddle, Woodside, NY, U.S.A.Books Puddle

    Venditore con 4 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 93,85

    EUR 3,55 spedizione 
    Spedito in U.S.A.

    Quantità: 4 disponibili

    Condizione: New.

  • Condizione: Nuovo

    EUR 84,12

    EUR 14,71 spedizione 
    Spedito da Regno Unito a U.S.A.

    Quantità: 2 disponibili

    Paperback. Condizione: Brand New. 2011 edition. 466 pages. 9.50x6.25x1.00 inches. In Stock.

  • Altre immagini

    Lingua: Inglese

    Editore: Springer, 2010

    303100423X / 9783031004230

    • Brossura

    Da: preigu, Osnabrück, Germaniapreigu

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 32,50

    EUR 70,00 spedizione 
    Spedito da Germania a U.S.A.

    Quantità: 5 disponibili

    Taschenbuch. Condizione: Neu. Algorithms for Reinforcement Learning | Csaba Szepesvári | Taschenbuch | Synthesis Lectures on Artificial Intelligence and Machine Learning | xiii | Englisch | 2010 | Springer | EAN 9783031004230 | Verantwortliche Person für die EU: Springer Verlag GmbH, Tiergartenstr. 17, 69121 Heidelberg, juergen[dot]hartmann[at]springer[dot]com | Anbieter: preigu. …

  • Lingua: Inglese

    Editore: Springer International Publishing AG, CH, 2010

    303100423X / 9783031004230

    • Brossura
    • Prima edizione

    Da: Rarewaves.com UK, London, Regno UnitoRarewaves.com UK

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 36,10

    EUR 76,48 spedizione 
    Spedito da Regno Unito a U.S.A.

    Quantità: Più di 20 disponibili

    Paperback. Condizione: New. 1st. Reinforcement learning is a learning paradigm concerned with learning to control a system so as to maximize a numerical performance measure that expresses a long-term objective. What distinguishes reinforcement learning from supervised learning is that only partial feedback is given to the learner about the learner's predictions. Further, the predictions may have long term effects through influencing the future state of the controlled system. Thus, time plays a special role. The goal in reinforcement learning is to develop efficient learning algorithms, as well as to understand the algorithms' merits and limitations. Reinforcement learning is of great interest because of the large number of practical applications that it can be used to address, ranging from problems in artificial intelligence to operations research or control engineering. In this book, we focus on those algorithms of reinforcement learning that build on the powerful theory of dynamic programming. We give a fairly comprehensive catalog of learning problems, describe the core ideas, note a large number of state of the art algorithms, followed by the discussion of their theoretical properties and limitations. Table of Contents: Markov Decision Processes / Value Prediction Problems / Control / For Further Exploration.…

  • Lingua: Inglese

    Editore: Berlin ; Heidelberg : Springer, 2011

    3642244114 / 9783642244117

    • Brossura

    Da: BBB-Internetbuchantiquariat, Bremen, GermaniaBBB-Internetbuchantiquariat

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Usato - Ottimo

    EUR 39,30

    EUR 79,00 spedizione 
    Spedito da Germania a U.S.A.

    Quantità: 1 disponibile

    Softcover/Paperback. Condizione: Sehr gut. 451 Seiten Zustand: sehr gut; Ungelesen; Fußschnitt leicht angeschmutzt; T-AA1357 9783642244117 Wenn das Buch einen Schutzumschlag hat, ist das ausdrücklich erwähnt. Rechnung mit ausgewiesener Mwst. Sprache: Englisch Gewicht in Gramm: 745.

  • Lingua: Inglese

    Editore: Springer-Verlag Berlin and Heidelberg GmbH & Co. KG, Berlin, 2011

    3642244114 / 9783642244117

    • Brossura

    Da: AussieBookSeller, Truganina, VIC, AustraliaAussieBookSeller

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 90,58

    EUR 32,88 spedizione 
    Spedito da Australia a U.S.A.

    Quantità: 1 disponibile

    Paperback. Condizione: new. Paperback. This book constitutes the refereed proceedings of the 22nd International Conference on Algorithmic Learning Theory, ALT 2011, held in Espoo, Finland, in October 2011, co-located with the 14th International Conference on Discovery Science, DS 2011. The 28 revised full papers presented together with the abstracts of 5 invited talks were carefully reviewed and selected from numerous submissions. The papers are divided into topical sections of papers on inductive inference, regression, bandit problems, online learning, kernel and margin-based methods, intelligent agents and other learning models. This book constitutes the refereed proceedings of the 22nd International Conference on Algorithmic Learning Theory, ALT 2011, held in Espoo, Finland, in October 2011, co-located with the 14th International Conference on Discovery Science, DS 2011. Shipping may be from our Sydney, NSW warehouse or from our UK or US warehouse, depending on stock availability.…

  • Lingua: Inglese

    Editore: Wiley, 2003

    0471498092 / 9780471498094

    • Rilegato

    Da: SMASS Sellers, IRVING, TX, U.S.A.SMASS Sellers

    Venditore con 4 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 125,33

     Spedizione gratuita 
    Spedito in U.S.A.

    Quantità: 5 disponibili

    Condizione: New. Brand New Original US Edition. Customer service! Satisfaction Guaranteed.

  • Lingua: Inglese

    Editore: Wiley, 2003

    0471498092 / 9780471498094

    • Rilegato

    Da: Romtrade Corp., STERLING HEIGHTS, MI, U.S.A.Romtrade Corp.

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 135,96

     Spedizione gratuita 
    Spedito in U.S.A.

    Quantità: 5 disponibili

    Condizione: New. This is a Brand-new US Edition. This Item may be shipped from US or any other country as we have multiple locations worldwide.

  • Lingua: Inglese

    Editore: Springer-Verlag Berlin and Heidelberg GmbH and Co. KG, DE, 2011

    3642244114 / 9783642244117

    • Brossura

    Da: Rarewaves.com UK, London, Regno UnitoRarewaves.com UK

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 67,58

    EUR 76,48 spedizione 
    Spedito da Regno Unito a U.S.A.

    Quantità: Più di 20 disponibili

    Paperback. Condizione: New. 2011th. This book constitutes the refereed proceedings of the 22nd International Conference on Algorithmic Learning Theory, ALT 2011, held in Espoo, Finland, in October 2011, co-located with the 14th International Conference on Discovery Science, DS 2011. The 28 revised full papers presented together with the abstracts of 5 invited talks were carefully reviewed and selected from numerous submissions. The papers are divided into topical sections of papers on inductive inference, regression, bandit problems, online learning, kernel and margin-based methods, intelligent agents and other learning models.…

  • Lingua: Inglese

    Editore: Cambridge University Press, GB, 2020

    1108486827 / 9781108486828

    • Rilegato

    Da: Rarewaves.com UK, London, Regno UnitoRarewaves.com UK

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Nuovo

    EUR 74,28

    EUR 76,48 spedizione 
    Spedito da Regno Unito a U.S.A.

    Quantità: Più di 20 disponibili

    Hardback. Condizione: New. Decision-making in the face of uncertainty is a significant challenge in machine learning, and the multi-armed bandit model is a commonly used framework to address it. This comprehensive and rigorous introduction to the multi-armed bandit problem examines all the major settings, including stochastic, adversarial, and Bayesian frameworks. A focus on both mathematical intuition and carefully worked proofs makes this an excellent reference for established researchers and a helpful resource for graduate students in computer science, engineering, statistics, applied mathematics and economics. Linear bandits receive special attention as one of the most useful models in applications, while other chapters are dedicated to combinatorial bandits, ranking, non-stationary problems, Thompson sampling and pure exploration. The book ends with a peek into the world beyond bandits with an introduction to partial monitoring and learning in Markov decision processes.…

  • Lingua: Inglese

    Editore: Springer-Verlag GmbH, 2011

    3642244114 / 9783642244117

    • Brossura

    Da: Buchpark, Trebbin, GermaniaBuchpark

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Usato - Ottimo

    EUR 43,93

    EUR 105,00 spedizione 
    Spedito da Germania a U.S.A.

    Quantità: 1 disponibile

    Condizione: Sehr gut. Zustand: Sehr gut | Seiten: 451 | Sprache: Englisch | Produktart: Bücher | Keine Beschreibung verfügbar.

  • Lingua: Inglese

    Editore: Morgan and Claypool Publishers, 2010

    1608454924 / 9781608454921

    • Brossura

    Da: YESIBOOKSTORE, MIAMI, FL, U.S.A.YESIBOOKSTORE

    Venditore con 5 stelle
    Contatta il venditore

    Condizione: Usato - Come nuovo

    EUR 166,49

     Spedizione gratuita 
    Spedito in U.S.A.

    Quantità: 1 disponibile

    paperback. Condizione: As New.