Csaba szepesvari (44 résultats)
- Couverture souple
Vendeur : World of Books (was SecondSale), Montgomery, IL, Etats-UnisWorld of Books (was SecondSale)
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Occasion - Satisfaisant
EUR 33,90
Frais de port gratuitsExpédition nationale : Etats-UnisQuantité disponible : 1 disponible(s)
Etat : Good. Item in good condition. Textbooks may not include supplemental items i.e. CDs, access codes etc.
Langue : anglais
Edité par Springer International Publishing AG, Cham, 2010
- Couverture souple
Vendeur : Grand Eagle Retail, Bensenville, IL, Etats-UnisGrand Eagle Retail
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 33,90
Frais de port gratuitsExpédition nationale : Etats-UnisQuantité disponible : 1 disponible(s)
Paperback. Etat : new. Paperback. Reinforcement learning is a learning paradigm concerned with learning to control a system so as to maximize a numerical performance measure that expresses a long-term objective. What distinguishes reinforcement learning from supervised learning is that only partial feedback is given to the learn…er about the learner's predictions. Further, the predictions may have long term effects through influencing the future state of the controlled system. Thus, time plays a special role. The goal in reinforcement learning is to develop efficient learning algorithms, as well as to understand the algorithms' merits and limitations. Reinforcement learning is of great interest because of the large number of practical applications that it can be used to address, ranging from problems in artificial intelligence to operations research or control engineering. In this book, we focus on those algorithms of reinforcement learning that build on the powerful theory of dynamic programming. We give a fairly comprehensive catalog of learning problems, describe the core ideas, note a large number of state of the art algorithms, followed by the discussion of their theoretical properties and limitations. Table of Contents: Markov Decision Processes / Value Prediction Problems / Control / For Further Exploration Reinforcement learning is a learning paradigm concerned with learning to control a system so as to maximize a numerical performance measure that expresses a long-term objective. What distinguishes reinforcement learning from supervised learning is that only partial feedback is given to the learner about the learner's predictions. Shipping may be from multiple locations in the US or from the UK, depending on stock availability.
- Couverture souple
Vendeur : California Books, Miami, FL, Etats-UnisCalifornia Books
Contacter le vendeurVendeur avec une évaluation de 4 étoilesEtat: Neuf
EUR 37,48
Frais de port gratuitsExpédition nationale : Etats-UnisQuantité disponible : Plus de 20 disponibles
Etat : New.
- Autres images
- Couverture souple
- Édition originale
Vendeur : Rarewaves.com USA, London, LONDO, Royaume-UniRarewaves.com USA
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 37,87
Frais de port gratuitsExpédition depuis Royaume-Uni vers Etats-UnisQuantité disponible : Plus de 20 disponibles
Paperback. Etat : New. 1st. Reinforcement learning is a learning paradigm concerned with learning to control a system so as to maximize a numerical performance measure that expresses a long-term objective. What distinguishes reinforcement learning from supervised learning is that only partial feedback is given to the learner abo…ut the learner's predictions. Further, the predictions may have long term effects through influencing the future state of the controlled system. Thus, time plays a special role. The goal in reinforcement learning is to develop efficient learning algorithms, as well as to understand the algorithms' merits and limitations. Reinforcement learning is of great interest because of the large number of practical applications that it can be used to address, ranging from problems in artificial intelligence to operations research or control engineering. In this book, we focus on those algorithms of reinforcement learning that build on the powerful theory of dynamic programming. We give a fairly comprehensive catalog of learning problems, describe the core ideas, note a large number of state of the art algorithms, followed by the discussion of their theoretical properties and limitations. Table of Contents: Markov Decision Processes / Value Prediction Problems / Control / For Further Exploration.
- Couverture souple
Vendeur : Books Puddle, New York, NY, Etats-UnisBooks Puddle
Contacter le vendeurVendeur avec une évaluation de 4 étoilesEtat: Neuf
EUR 35,78
EUR 3,46 expéditionExpédition nationale : Etats-UnisQuantité disponible : 1 disponible(s)
Etat : New.
- Couverture souple
Vendeur : Romtrade Corp., STERLING HEIGHTS, MI, Etats-UnisRomtrade Corp.
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 42,32
Frais de port gratuitsExpédition nationale : Etats-UnisQuantité disponible : 1 disponible(s)
Etat : New. This is a Brand-new US Edition. This Item may be shipped from US or any other country as we have multiple locations worldwide.
- Couverture souple
Vendeur : Basi6 International, Irving, TX, Etats-UnisBasi6 International
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 42,32
Frais de port gratuitsExpédition nationale : Etats-UnisQuantité disponible : 3 disponible(s)
Etat : Brand New. New. US edition. Expediting shipping for all USA and Europe orders excluding PO Box. Excellent Customer Service.
- Couverture souple
Vendeur : SMASS Sellers, IRVING, TX, Etats-UnisSMASS Sellers
Contacter le vendeurVendeur avec une évaluation de 4 étoilesEtat: Neuf
EUR 43,80
Frais de port gratuitsExpédition nationale : Etats-UnisQuantité disponible : 1 disponible(s)
Etat : New. Brand New Original US Edition. Customer service! Satisfaction Guaranteed.
- Couverture souple
Vendeur : Ria Christie Collections, Uxbridge, Royaume-UniRia Christie Collections
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 34,75
EUR 14,00 expéditionExpédition depuis Royaume-Uni vers Etats-UnisQuantité disponible : Plus de 20 disponibles
Etat : New. In English.
Langue : anglais
Edité par Springer-Verlag Berlin and Heidelberg GmbH & Co. KG, Berlin, 2011
- Couverture souple
Vendeur : Grand Eagle Retail, Bensenville, IL, Etats-UnisGrand Eagle Retail
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 55,94
Frais de port gratuitsExpédition nationale : Etats-UnisQuantité disponible : 1 disponible(s)
Paperback. Etat : new. Paperback. This book constitutes the refereed proceedings of the 22nd International Conference on Algorithmic Learning Theory, ALT 2011, held in Espoo, Finland, in October 2011, co-located with the 14th International Conference on Discovery Science, DS 2011. The 28 revised full papers presented together wi…th the abstracts of 5 invited talks were carefully reviewed and selected from numerous submissions. The papers are divided into topical sections of papers on inductive inference, regression, bandit problems, online learning, kernel and margin-based methods, intelligent agents and other learning models. This book constitutes the refereed proceedings of the 22nd International Conference on Algorithmic Learning Theory, ALT 2011, held in Espoo, Finland, in October 2011, co-located with the 14th International Conference on Discovery Science, DS 2011. Shipping may be from multiple locations in the US or from the UK, depending on stock availability.
- Couverture rigide
Vendeur : California Books, Miami, FL, Etats-UnisCalifornia Books
Contacter le vendeurVendeur avec une évaluation de 4 étoilesEtat: Neuf
EUR 61,57
Frais de port gratuitsExpédition nationale : Etats-UnisQuantité disponible : Plus de 20 disponibles
Etat : New.
- Autres images
Langue : anglais
Edité par Springer-Verlag Berlin and Heidelberg GmbH and Co. KG, DE, 2011
- Couverture souple
Vendeur : Rarewaves.com USA, London, LONDO, Royaume-UniRarewaves.com USA
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 69,50
Frais de port gratuitsExpédition depuis Royaume-Uni vers Etats-UnisQuantité disponible : Plus de 20 disponibles
Paperback. Etat : New. 2011th. This book constitutes the refereed proceedings of the 22nd International Conference on Algorithmic Learning Theory, ALT 2011, held in Espoo, Finland, in October 2011, co-located with the 14th International Conference on Discovery Science, DS 2011. The 28 revised full papers presented together with…the abstracts of 5 invited talks were carefully reviewed and selected from numerous submissions. The papers are divided into topical sections of papers on inductive inference, regression, bandit problems, online learning, kernel and margin-based methods, intelligent agents and other learning models.
- Couverture rigide
Vendeur : Books Puddle, New York, NY, Etats-UnisBooks Puddle
Contacter le vendeurVendeur avec une évaluation de 4 étoilesEtat: Neuf
EUR 67,24
EUR 3,46 expéditionExpédition nationale : Etats-UnisQuantité disponible : 4 disponible(s)
Etat : New.
- Couverture souple
Vendeur : Ria Christie Collections, Uxbridge, Royaume-UniRia Christie Collections
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 61,15
EUR 14,00 expéditionExpédition depuis Royaume-Uni vers Etats-UnisQuantité disponible : Plus de 20 disponibles
Etat : New. In.
Algorithmic Learning Theory: 22nd International Conference, ALT 2011, Espoo, Finland, October 5-7, 2011, Proceedings (Lecture Notes in Computer Science)
Jyriki Kivinen, Csaba Szepesv�ri, Esko Ukkonen, Thomas Zeugmann
- Couverture souple
Vendeur : Chiron Media, Wallingford, Royaume-UniChiron Media
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 57,75
EUR 18,10 expéditionExpédition depuis Royaume-Uni vers Etats-UnisQuantité disponible : 10 disponible(s)
Paperback. Etat : New.
- Couverture rigide
Vendeur : Ria Christie Collections, Uxbridge, Royaume-UniRia Christie Collections
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 62,39
EUR 14,00 expéditionExpédition depuis Royaume-Uni vers Etats-UnisQuantité disponible : Plus de 20 disponibles
Etat : New. In.
- Autres images
- Couverture rigide
Vendeur : Rarewaves.com USA, London, LONDO, Royaume-UniRarewaves.com USA
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 77,85
Frais de port gratuitsExpédition depuis Royaume-Uni vers Etats-UnisQuantité disponible : 1 disponible(s)
Hardback. Etat : New. Decision-making in the face of uncertainty is a significant challenge in machine learning, and the multi-armed bandit model is a commonly used framework to address it. This comprehensive and rigorous introduction to the multi-armed bandit problem examines all the major settings, including stochastic, advers…arial, and Bayesian frameworks. A focus on both mathematical intuition and carefully worked proofs makes this an excellent reference for established researchers and a helpful resource for graduate students in computer science, engineering, statistics, applied mathematics and economics. Linear bandits receive special attention as one of the most useful models in applications, while other chapters are dedicated to combinatorial bandits, ranking, non-stationary problems, Thompson sampling and pure exploration. The book ends with a peek into the world beyond bandits with an introduction to partial monitoring and learning in Markov decision processes.
Langue : anglais
Edité par Springer International Publishing AG, Cham, 2010
- Couverture souple
Vendeur : AussieBookSeller, Truganina, VIC, AustralieAussieBookSeller
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 52,87
EUR 32,05 expéditionExpédition depuis Australie vers Etats-UnisQuantité disponible : 1 disponible(s)
Paperback. Etat : new. Paperback. Reinforcement learning is a learning paradigm concerned with learning to control a system so as to maximize a numerical performance measure that expresses a long-term objective. What distinguishes reinforcement learning from supervised learning is that only partial feedback is given to the learn…er about the learner's predictions. Further, the predictions may have long term effects through influencing the future state of the controlled system. Thus, time plays a special role. The goal in reinforcement learning is to develop efficient learning algorithms, as well as to understand the algorithms' merits and limitations. Reinforcement learning is of great interest because of the large number of practical applications that it can be used to address, ranging from problems in artificial intelligence to operations research or control engineering. In this book, we focus on those algorithms of reinforcement learning that build on the powerful theory of dynamic programming. We give a fairly comprehensive catalog of learning problems, describe the core ideas, note a large number of state of the art algorithms, followed by the discussion of their theoretical properties and limitations. Table of Contents: Markov Decision Processes / Value Prediction Problems / Control / For Further Exploration Reinforcement learning is a learning paradigm concerned with learning to control a system so as to maximize a numerical performance measure that expresses a long-term objective. What distinguishes reinforcement learning from supervised learning is that only partial feedback is given to the learner about the learner's predictions. Shipping may be from our Sydney, NSW warehouse or from our UK or US warehouse, depending on stock availability.
- Autres images
- Couverture souple
Vendeur : AHA-BUCH GmbH, Einbeck, AllemagneAHA-BUCH GmbH
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 32,09
EUR 61,06 expéditionExpédition depuis Allemagne vers Etats-UnisQuantité disponible : 1 disponible(s)
Taschenbuch. Etat : Neu. Druck auf Anfrage Neuware - Printed after ordering - Reinforcement learning is a learning paradigm concerned with learning to control a system so as to maximize a numerical performance measure that expresses a long-term objective. What distinguishes reinforcement learning from supervised learning is that… only partial feedback is given to the learner about the learner's predictions. Further, the predictions may have long term effects through influencing the future state of the controlled system. Thus, time plays a special role. The goal in reinforcement learning is to develop efficient learning algorithms, as well as to understand the algorithms' merits and limitations. Reinforcement learning is of great interest because of the large number of practical applications that it can be used to address, ranging from problems in artificial intelligence to operations research or control engineering. In this book, we focus on those algorithms of reinforcement learning that build on the powerful theory of dynamic programming. We give a fairly comprehensive catalog of learning problems, describe the core ideas, note a large number of state of the art algorithms, followed by the discussion of their theoretical properties and limitations. Table of Contents: Markov Decision Processes / Value Prediction Problems / Control / For Further Exploration.
- Couverture souple
Vendeur : Revaluation Books, Exeter, Royaume-UniRevaluation Books
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 82,03
EUR 14,61 expéditionExpédition depuis Royaume-Uni vers Etats-UnisQuantité disponible : 2 disponible(s)
Paperback. Etat : Brand New. 2011 edition. 466 pages. 9.50x6.25x1.00 inches. In Stock.
- Autres images
- Couverture souple
Vendeur : preigu, Osnabrück, Allemagnepreigu
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 32,45
EUR 70,00 expéditionExpédition depuis Allemagne vers Etats-UnisQuantité disponible : 5 disponible(s)
Taschenbuch. Etat : Neu. Algorithms for Reinforcement Learning | Csaba Szepesvári | Taschenbuch | Synthesis Lectures on Artificial Intelligence and Machine Learning | xiii | Englisch | 2010 | Springer | EAN 9783031004230 | Verantwortliche Person für die EU: Springer Verlag GmbH, Tiergartenstr. 17, 69121 Heidelberg, juergen[dot]h…artmann[at]springer[dot]com | Anbieter: preigu.
- Autres images
- Couverture souple
- Édition originale
Vendeur : Rarewaves.com UK, London, Royaume-UniRarewaves.com UK
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 34,65
EUR 75,97 expéditionExpédition depuis Royaume-Uni vers Etats-UnisQuantité disponible : Plus de 20 disponibles
Paperback. Etat : New. 1st.
- Couverture rigide
Vendeur : Romtrade Corp., STERLING HEIGHTS, MI, Etats-UnisRomtrade Corp.
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 116,52
Frais de port gratuitsExpédition nationale : Etats-UnisQuantité disponible : 5 disponible(s)
Etat : New. This is a Brand-new US Edition. This Item may be shipped from US or any other country as we have multiple locations worldwide.
- Autres images
- Couverture rigide
Vendeur : moluna, Greven, Allemagnemoluna
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 64,81
EUR 48,99 expéditionExpédition depuis Allemagne vers Etats-UnisQuantité disponible : 2 disponible(s)
Etat : New. Decision-making in the face of uncertainty is a challenge in machine learning, and the multi-armed bandit model is a common framework to address it. This comprehensive introduction is an excellent reference for established researchers and a resource for gra.
Algorithmic learning theory 22nd international conference ; proceedings.
Kivinen, Jyrki , Csaba Szepesvári and Esko Zeugmann Thomas [Hrsg.] Ukkonen:
- Couverture souple
Vendeur : BBB-Internetbuchantiquariat, Bremen, AllemagneBBB-Internetbuchantiquariat
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Occasion - Très bon
EUR 39,30
EUR 79,00 expéditionExpédition depuis Allemagne vers Etats-UnisQuantité disponible : 1 disponible(s)
Softcover/Paperback. Etat : Sehr gut. 451 Seiten Zustand: sehr gut; Ungelesen; Fußschnitt leicht angeschmutzt; T-AA1357 9783642244117 Wenn das Buch einen Schutzumschlag hat, ist das ausdrücklich erwähnt. Rechnung mit ausgewiesener Mwst. Sprache: Englisch Gewicht in Gramm: 745.
- Couverture rigide
Vendeur : SMASS Sellers, IRVING, TX, Etats-UnisSMASS Sellers
Contacter le vendeurVendeur avec une évaluation de 4 étoilesEtat: Neuf
EUR 122,17
Frais de port gratuitsExpédition nationale : Etats-UnisQuantité disponible : 5 disponible(s)
Etat : New. Brand New Original US Edition. Customer service! Satisfaction Guaranteed.
Langue : anglais
Edité par Springer-Verlag Berlin and Heidelberg GmbH & Co. KG, Berlin, 2011
- Couverture souple
Vendeur : AussieBookSeller, Truganina, VIC, AustralieAussieBookSeller
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 92,69
EUR 32,05 expéditionExpédition depuis Australie vers Etats-UnisQuantité disponible : 1 disponible(s)
Paperback. Etat : new. Paperback. This book constitutes the refereed proceedings of the 22nd International Conference on Algorithmic Learning Theory, ALT 2011, held in Espoo, Finland, in October 2011, co-located with the 14th International Conference on Discovery Science, DS 2011. The 28 revised full papers presented together wi…th the abstracts of 5 invited talks were carefully reviewed and selected from numerous submissions. The papers are divided into topical sections of papers on inductive inference, regression, bandit problems, online learning, kernel and margin-based methods, intelligent agents and other learning models. This book constitutes the refereed proceedings of the 22nd International Conference on Algorithmic Learning Theory, ALT 2011, held in Espoo, Finland, in October 2011, co-located with the 14th International Conference on Discovery Science, DS 2011. Shipping may be from our Sydney, NSW warehouse or from our UK or US warehouse, depending on stock availability.
- Autres images
- Couverture rigide
Vendeur : AHA-BUCH GmbH, Einbeck, AllemagneAHA-BUCH GmbH
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 68,24
EUR 65,14 expéditionExpédition depuis Allemagne vers Etats-UnisQuantité disponible : 2 disponible(s)
Buch. Etat : Neu. Neuware.
Algorithmic Learning Theory: 22nd International Conference, ALT 2011, Espoo, Finland, October 5-7, 2011, Proceedings: 6925 (Lecture Notes in Computer Science, 6925)
Jyriki Kivinen, Csaba Szepesvári, Esko Ukkonen, Thomas Zeugmann
- Couverture souple
Vendeur : Rarewaves.com UK, London, Royaume-UniRarewaves.com UK
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 64,78
EUR 75,97 expéditionExpédition depuis Royaume-Uni vers Etats-UnisQuantité disponible : Plus de 20 disponibles
Paperback. Etat : New. 2011th.
- Autres images
- Couverture rigide
Vendeur : preigu, Osnabrück, Allemagnepreigu
Contacter le vendeurVendeur avec une évaluation de 5 étoilesEtat: Neuf
EUR 72,65
EUR 70,00 expéditionExpédition depuis Allemagne vers Etats-UnisQuantité disponible : 1 disponible(s)
Buch. Etat : Neu. Bandit Algorithms | Csaba Szepesvari (u. a.) | Buch | Gebunden | Englisch | 2020 | Cambridge University Press | EAN 9781108486828 | Verantwortliche Person für die EU: Libri GmbH, Europaallee 1, 36244 Bad Hersfeld, gpsr[at]libri[dot]de | Anbieter: preigu.












