The Art of Reinforcement Learning

預購

台灣風情茄芷袋Supercard造型悠遊卡-台灣(裁型)

可愛ｘ實用ｘ好旅伴，台味悠遊卡熱銷中！

喜歡+1
寫評價
賺金幣

9折 2052元
~~2280~~元
認購希望書包，幫助弱勢孩童上學不中斷！

預計最高可得金幣100點 ? 可100%折抵
活動加倍另計
HAPPY GO享100累1點 4點抵1元 折抵無上限

分類：
英文書＞自然科普＞電腦資訊＞其他電腦資訊
追蹤

? 追蹤分類後，您會在第一時間收到分類新品通知。
作者： Michael,Hu 追蹤 ? 追蹤作者後，您會在第一時間收到作者新書通知。
出版社： Apress 追蹤 ? 追蹤出版社後，您會在第一時間收到出版社新書通知。
出版日：2023/12/09

信用卡分期： 6期 0利率每期 342元更多分期

分期價：除不盡餘數於第一期收取
3期0利率	每期684 元	接受26 家銀行
6期0利率	每期342 元	接受26 家銀行

3期0利率接受26家銀行

土地銀行、合作金庫、第一銀行、華南銀行、上海銀行、台北富邦、兆豐商銀、花旗(台灣)銀行、澳盛銀行、臺灣企銀、渣打商銀、滙豐(台灣)銀行、臺灣新光商銀、陽信銀行、三信銀行、聯邦銀行、遠東銀行、元大銀行、永豐銀行、玉山銀行、星展銀行、台新銀行、日盛銀行、安泰銀行、中國信託、台灣樂天

6期0利率接受26家銀行

12期0利率接受26家銀行

24期0利率接受22家銀行

土地銀行、合作金庫、第一銀行、華南銀行、上海銀行、台北富邦、花旗(台灣)銀行、澳盛銀行、臺灣企銀、渣打商銀、滙豐(台灣)銀行、臺灣新光商銀、陽信銀行、聯邦銀行、遠東銀行、元大銀行、玉山銀行、星展銀行、台新銀行、日盛銀行、安泰銀行、中國信託

立即代訂

※此為代訂海外書籍，不接受退換貨

※ 本商品會員日滿額金幣加碼回饋最高15倍

購買後進貨　

預訂門市商品

門市庫存

內容簡介

Unlock the full potential of reinforcement learning (RL), a crucial subfield of Artificial Intelligence, with this comprehensive guide. This book provides a deep dive into RL's core concepts, mathematics, and practical algorithms, helping you to develop a thorough understanding of this cutting-edge technology. Beginning with an overview of fundamental concepts such as Markov decision processes, dynamic programming, Monte Carlo methods, and temporal difference learning, this book uses clear and concise examples to explain the basics of RL theory. The following section covers value function approximation, a critical technique in RL, and explores various policy approximations such as policy gradient methods and advanced algorithms like Proximal Policy Optimization (PPO).

This book also delves into advanced topics, including distributed reinforcement learning, curiosity-driven exploration, and the famous AlphaZero algorithm, providing readers with a detailed account of these cutting-edge techniques.

With a focus on explaining algorithms and the intuition behind them, The Art of Reinforcement Learning includes practical source code examples that you can use to implement RL algorithms. Upon completing this book, you will have a deep understanding of the concepts, mathematics, and algorithms behind reinforcement learning, making it an essential resource for AI practitioners, researchers, and students.

What You Will Learn

Grasp fundamental concepts and distinguishing features of reinforcement learning, including how it differs from other AI and non-interactive machine learning approachesModel problems as Markov decision processes, and how to evaluate and optimize policies using dynamic programming, Monte Carlo methods, and temporal difference learningUtilize techniques for approximating value functions and policies, including linear and nonlinear value function approximation and policy gradient methodsUnderstand the architecture and advantages of distributed reinforcement learningMaster the concept of curiosity-driven exploration and how it can be leveraged to improve reinforcement learning agentsExplore the AlphaZero algorithm and how it was able to beat professional Go players

Who This Book Is For

Machine learning engineers, data scientists, software engineers, and developers who want to incorporate reinforcement learning algorithms into their projects and applications.

配送方式

台灣
- 國內宅配：本島、離島
- 到店取貨：
  
  不限金額免運費
海外
- 國際快遞：全球
- 港澳店取：

詳細資料

- 語言
- 英文
- 裝訂
- 紙本平裝
- ISBN
- 9781484296059
- 分級
- 普通級
- 頁數
- 0
- 商品規格
- 出版地
- 美國
- 適讀年齡
- 全齡適讀
- 注音
- 級別

英文書＞自然科普＞電腦資訊＞其他電腦資訊

商品評價

訂購/退換貨須知

加入金石堂 LINE 官方帳號『完成綁定』，隨時掌握出貨動態：

商品運送說明：

本公司所提供的產品配送區域範圍目前僅限台灣本島。注意！收件地址請勿為郵政信箱。
商品將由廠商透過貨運或是郵局寄送。消費者訂購之商品若無法送達，經電話或 E-mail無法聯繫逾三天者，本公司將取消該筆訂單，並且全額退款。
當廠商出貨後，您會收到E-mail出貨通知，您也可透過【訂單查詢】確認出貨情況。
產品顏色可能會因網頁呈現與拍攝關係產生色差，圖片僅供參考，商品依實際供貨樣式為準。
如果是大型商品（如：傢俱、床墊、家電、運動器材等）及需安裝商品，請依商品頁面說明為主。訂單完成收款確認後，出貨廠商將會和您聯繫確認相關配送等細節。
偏遠地區、樓層費及其它加價費用，皆由廠商於約定配送時一併告知，廠商將保留出貨與否的權利。

提醒您！！
金石堂及銀行均不會請您操作ATM! 如接獲電話要求您前往ATM提款機，請不要聽從指示，以免受騙上當！

退換貨須知：

**提醒您，鑑賞期不等於試用期，退回商品須為全新狀態**

依據「消費者保護法」第19條及行政院消費者保護處公告之「通訊交易解除權合理例外情事適用準則」，以下商品購買後，除商品本身有瑕疵外，將不提供7天的猶豫期：
1. 易於腐敗、保存期限較短或解約時即將逾期。（如：生鮮食品）
2. 依消費者要求所為之客製化給付。（客製化商品）
3. 報紙、期刊或雜誌。（含MOOK、外文雜誌）
4. 經消費者拆封之影音商品或電腦軟體。
5. 非以有形媒介提供之數位內容或一經提供即為完成之線上服務，經消費者事先同意始提供。（如：電子書、電子雜誌、下載版軟體、虛擬商品…等）
6. 已拆封之個人衛生用品。（如：內衣褲、刮鬍刀、除毛刀…等）
若非上列種類商品，均享有到貨7天的猶豫期（含例假日）。
辦理退換貨時，商品（組合商品恕無法接受單獨退貨）必須是您收到商品時的原始狀態（包含商品本體、配件、贈品、保證書、所有附隨資料文件及原廠內外包裝…等），請勿直接使用原廠包裝寄送，或於原廠包裝上黏貼紙張或書寫文字。
退回商品若無法回復原狀，將請您負擔回復原狀所需費用，嚴重時將影響您的退貨權益。