Your first multi-agent RL project will teach you a hard truth: everything you know about single-agent training breaks the moment a second learner enters the room.
Non-stationarity sets in. Rewards stop meaning what you think they mean. And the fixes that worked for a single policy quietly make things worse.
This is the book for engineers and researchers who already know single-agent RL and are ready for what comes next — written by a practitioner who's built coordinating robot fleets, adversarial trading agents, and cooperating LLM agent teams, and who still remembers exactly where it went wrong the first time.
Inside, you'll learn:
Les informations fournies dans la section « Synopsis » peuvent faire référence à une autre édition de ce titre.
Vendeur : California Books, Miami, FL, Etats-Unis
Etat : New. Print on Demand. N° de réf. du vendeur I-9798185265246
Quantité disponible : Plus de 20 disponibles
Vendeur : PBShop.store US, Wood Dale, IL, Etats-Unis
PAP. Etat : New. New Book. Shipped from UK. Established seller since 2000. N° de réf. du vendeur L2-9798185265246
Quantité disponible : Plus de 20 disponibles
Vendeur : PBShop.store UK, Fairford, GLOS, Royaume-Uni
PAP. Etat : New. New Book. Shipped from UK. Established seller since 2000. N° de réf. du vendeur L2-9798185265246
Quantité disponible : Plus de 20 disponibles
Vendeur : AHA-BUCH GmbH, Einbeck, Allemagne
Taschenbuch. Etat : Neu. Neuware. N° de réf. du vendeur 9798185265246
Quantité disponible : 2 disponible(s)