Reinforcement Learning: Theory and Python Implementation

Author:   Zhiqing Xiao
Publisher:   Springer
ISBN:  

9789811949357


Pages:   559
Publication Date:   30 September 2025
Format:   Paperback
Availability:   Manufactured on demand   Availability explained
We will order this item for you from a manufactured on demand supplier.

Our Price $158.37 Quantity:  
Add to Cart

Share |

Reinforcement Learning: Theory and Python Implementation


Overview

Reinforcement Learning: Theory and Python Implementation is a tutorial book on reinforcement learning, with explanations of both theory and applications. Starting from a uniform mathematical framework, this book derives the theory of modern reinforcement learning systematically and introduces all mainstream reinforcement learning algorithms such as PPO, SAC, and MuZero. It also covers key technologies of GPT training such as RLHF, IRL, and PbRL. Every chapter is accompanied by high-quality implementations, and all implementations of deep reinforcement learning algorithms are with both TensorFlow and PyTorch. Codes can be found on GitHub along with their results and are runnable on a conventional laptop with either Windows, macOS, or Linux. This book is intended for readers who want to learn reinforcement learning systematically and apply reinforcement learning to practical applications. It is also ideal to academical researchers who seek theoretical foundation or algorithm enhancement in their cutting-edge AI research.

Full Product Details

Author:   Zhiqing Xiao
Publisher:   Springer
Imprint:   Springer
ISBN:  

9789811949357


ISBN 10:   9811949352
Pages:   559
Publication Date:   30 September 2025
Audience:   General/trade ,  General
Format:   Paperback
Publisher's Status:   Active
Availability:   Manufactured on demand   Availability explained
We will order this item for you from a manufactured on demand supplier.

Table of Contents

Chapter 1. Introduction of Reinforcement Learning (RL).- Chapter 2. MDP: Markov Decision Process.- Chapter 3. Model-based Numerical Iteration.- Chapter 4. MC: Monte Carlo Learning.- Chapter 5. TD: Temporal Difference Learning.- Chapter 6. Function Approximation.- Chapter 7. PG: Policy Gradient.- Chapter 8. AC: Actor–Critic.- Chapter 9. DPG: Deterministic Policy Gradient.- Chapter 10. Maximum-Entropy RL.- Chapter 11. Policy-based Gradient-Free Algorithms.- Chapter 12. Distributional RL.- Chapter 13. Minimize Regret.- Chapter 14. Tree Search.- Chapter 15. More Agent–Environment Interfaces.- Chapter 16. Learn from Feedback and Imitation Learning.

Reviews

“The book is an excellent resource for anyone looking to explore the world of reinforcement learning (RL). This book combines theoretical depth with practical implementation, making it a standout choice for students, researchers, and industry professionals alike. ... The book is a comprehensive guide that balances theoretical rigor with practical usability.” (Catalin Stoean, zbMATH 1562.68008, 2025)


Author Information

Zhiqing Xiao obtained doctoral degree from Tsinghua University in 2016 and has more than 15 years in academic research and industrial practices on data-analytics and AI. He is the author of two AI bestsellers in Chinese: “Reinforcement Learning” and “Application of Neural Network and PyTorch” and published many academic papers. He also contributed to recent versions of the open-source software Gym.

Tab Content 6

Author Website:  

Countries Available

All regions
Latest Reading Guide

NOV RG 20252

 

Shopping Cart
Your cart is empty
Shopping cart
Mailing List