Reinforcement Learning

Reinforcement Learning

By Trilokesh Khatri
Michael Caine
Listen with Sir Michael Caine™ and 1,000+ voices
Length7h 56m

About this audiobook

Reinforcement Learning: A Practical Guide to Algorithms delves into the impactful world of reinforcement learning, a key branch of AI. Spanning over five decades, reinforcement learning has significantly advanced AI, offering solutions for planning, budgeting, and strategic decision-making. This book provides a comprehensive understanding of reinforcement learning, focusing on building smart models and agents that adapt to changing requirements. We cover fundamental and advanced topics, including value-based methods like UCB, SARSA, and Q-learning, as well as function approximation techniques. Additionally, we explore artificial neural networks, LSTD, gradient methods, emphatic TD methods, average reward methods, and policy gradient methods. With clear explanations, diagrams, and examples, this book ensures that readers can grasp and apply reinforcement learning algorithms to real-world problems effectively. By the end, you will have a solid foundation in both theoretical and practical aspects of reinforcement learning.

Audiobook details

GenreTechnology, Science and Nature
Length7 hrs 56 mins
Narrated byListen with 1,000+ voices
FormateBook with Audio
Publish dateJan 3, 2025
LanguageEnglish

Table of contents

1Part-1
2Tabular Solution Methods
3Part-2
4Approximate Solution Methods
5Chapter 1. Introduction
Show all chapters
61.1 Reinforcement Learning
71.2 Examples
81.3 Elements of RL
91.4 Applications of RL
101.5 Summary
111.6 Questions
12Chapter 2. Multi-arm Bandits
132.1 An n-armed bandit problem
142.2 Action-value methods
152.3 Incremental implementation
162.4 Tracking a nonstationary problem
172.5 Optimistic Initial Values
182.6 Upper-Confidence-Bound-Action Selection
192.7 Gradient Bandit Algorithms
202.8 Associative Search (Contextual Bandits)
212.9 Summary
222.10 Questions
23Chapter 3. Solving Problems with Dynamic Programming
243.1 MDP
253.2 Categorizing RL algorithms
263.3 Dynamic Programming
273.4 Summary
283.5 Questions
29Chapter 4. Monte Carlo Methods
304.1 Monte Carlo prediction
314.2 Monte Carlo estimation of action values
324.3 Monte Carlo Control
334.4 Monte Carlo Control without Exploring
34Starts
354.5 Off policy Prediction via Importance Sampling
364.6 Incremental Implementation
374.7 Off-policy MC Control
384.8 Discounting-aware Importance sampling
394.9 Per-decision Importance Sampling
404.10 Summary
414.11 Questions
42Chapter 5. Temporal-Difference Learning
435.1 TD Prediction
445.2 Advantages of TD Prediction methods
455.3 Optimality of TD(0)
465.4 SARSA: On-policy TD Control
475.5 Q-learning: off-policy TD Control
485.6 Expected SARSA
495.7 Maximization Bias and Double Learning
505.7 Summary

More from Trilokesh Khatri

Frosty: The incredible true story of the boy from Doonside who became a Bathurst king
Frosty: The incredible true story of the boy from Doonside who became a Bathurst kingMark Winterbottom8h 9m4.8 (50)$26 · $0.00
Cadillac Desert, Revised and Updated Edition
Cadillac Desert, Revised and Updated EditionMarc Reisner27h 56m4.4 (13.6K)$30
The Honey Bus
The Honey BusMeredith May9h 31m4.4 (16.6K)$25 · $0.00
Operation Pedestal
Operation PedestalMax Hastings12h 29m4.4 (5.2K)$29 · $0.00
When the Heavens Went on Sale
When the Heavens Went on SaleAshlee Vance18h 20m4.4 (3K)$40 · $0.00
Everybody Has a Podcast (Except You)
Everybody Has a Podcast (Except You)Justin McElroy, Travis McElroy, Griffin McElroy5h 9m4.3 (2.3K)$24 · $0.00
Hands of Time
Hands of TimeRebecca Struthers8h 8m4.4 (1.3K)$26 · $0.00
Never Lost Again
Never Lost AgainBill Kilday10h 1m4.4 (805)$29 · $0.00
How the Internet Happened
How the Internet HappenedBrian McCullough13h 28m4.3 (2.7K)$23 · $0.00
Collected Writings of Nikola Tesla
Collected Writings of Nikola TeslaNikola Tesla, Thomas Commerford Martin21h 50m4.3 (1.7K)$1 · $0.00
Just Aspire
Just AspireAjai Chowdhry9h 13m4.4 (260)$29 · $0.00
The Little Book of Aliens
The Little Book of AliensAdam Frank8h 20m4.3 (1.3K)$26 · $0.00
The inventions, researches and writings of Nikola Tesla (Annotated)
The inventions, researches and writings of Nikola Tesla (Annotated)Thomas Commerford Martin16h 59m4.3 (1.3K)$2 · $0.00
Kargil
KargilV.P. Malik15h 18m4.3 (2K)$24 · $0.00
Experiments with Alternating Currents
Experiments with Alternating CurrentsNikola Tesla9h 6m4.3 (874)$1 · $0.00
Fire on the Horizon
Fire on the HorizonTom Shroder, John Konrad8h 23m4.3 (858)$26 · $0.00
The Smell of Kerosene (Annotated)
The Smell of Kerosene (Annotated)National Aeronautics and Space Administration, Donald L. Mallick, Peter W. Merlin11h 32m4.3 (523)$2 · $0.00
The Boy Who Harnessed the Wind
The Boy Who Harnessed the WindWilliam Kamkwamba, Bryan Mealer10h 5m4.2 (78.9K)$29 · $0.00
Mars Rover Curiosity
Mars Rover CuriosityRob Manning, William L. Simon7h 43m4.3 (1.7K)$20
Across the Airless Wilds
Across the Airless WildsEarl Swift10h 8m4.3 (1.3K)$29 · $0.00

You may also like

12 Bytes
12 BytesJeanette Winterson9h4.0 (2.5K)$20
Autonomy
AutonomyLawrence D. Burns, Christopher Shulgan11h 21m4.1 (1.1K)$29 · $0.00
ITIL Foundation Essentials ITIL 4 Edition - The ultimate revision guide, second edition
ITIL Foundation Essentials ITIL 4 Edition - The ultimate revision guide, second editionClaire Agutter1h 12m4.2 (462)$27 · $0.00
Day Trading Attention
Day Trading AttentionGary Vaynerchuk8h 25m4.3 (1.8K)$28 · $0.00
The Highway Code UK
The Highway Code UKDepartment for Transport3h 33m4.2 (36.6K)$11 · $0.00
The In-Car-Nation Code
The In-Car-Nation CodeDr. Engelbert Wimmer20h 15m$21 · $0.00
Artificial Intelligence
Artificial IntelligenceJ.C.Lesko2h 43m$15 · $0.00
Replugged
RepluggedSatnam Bains13h 47m$14 · $0.00
Disturbance rejection control for bipedal robot walkers
Disturbance rejection control for bipedal robot walkersJaime Arcos Legarda4h 24m$7 · $0.00
Master Ai
Master AiHenrique Xavier Oliveira2h 5m$6 · $0.00
Chatgpt Complete Guide
Chatgpt Complete GuideJoaquin Gener1h 23m$6 · $0.00
The 9 AI Powers: How to Build Digital Wealth in 2025
The 9 AI Powers: How to Build Digital Wealth in 2025Yasin Ali 1h 41m4.7 (12) Only with Ultra
The AI Revolution
The AI RevolutionTorsten J. Koerting3h 11m$5 · $0.00
The Artificial Author
The Artificial AuthorSimone Aliprandi7h 42m$15 · $0.00
The Evolution of Technology
The Evolution of Technologythe Speech Resource Company9h 30m$20
The Future Ready Organization
The Future Ready OrganizationGyan Nagpal7h 44m4.3 (35)$26 · $0.00
Me, My Customer, and AI
Me, My Customer, and AIHenrik Werdelin, Nicholas Thorne3h 41m3.3 (16)$20
Driverless
DriverlessHod Lipson, Melba Kurman9h 57m3.9 (552)$20
The Electric Slide
The Electric SlidePacky McCormick4h 31m5.0 (12)Free
Giant Brains, or, Machines That Think
Giant Brains, or, Machines That ThinkEdmund Callis Berkeley9h 21m$2.30