返回首页
苏宁会员
购物车 0
易付宝
手机苏宁

服务体验

店铺评分与同行业相比

用户评价:----

物流时效:----

售后服务:----

  • 服务承诺: 正品保障
  • 公司名称:
  • 所 在 地:

  • 正版新书]阿尔法零对最优模型预测自适应控制的启示[美]德梅萃·P
  • 全店均为全新正版书籍,欢迎选购!新疆西藏青海(可包挂刷).港澳台及海外地区bu bao快递
    • 作者: [美]德梅萃·P. 博塞克斯(Dimitri P. Bertsekas) 著著 | [美]德梅萃·P. 博塞克斯(Dimitri P. Bertsekas) 著编 | [美]德梅萃·P. 博塞克斯(Dimitri P. Bertsekas) 著译 | [美]德梅萃·P. 博塞克斯(Dimitri P. Bertsekas) 著绘
    • 出版社: 清华大学出版社
    • 出版时间:2025-04-01
    送至
  • 由""直接销售和发货,并提供售后服务
  • 加入购物车 购买电子书
    服务

    看了又看

    商品预定流程:

    查看大图
    /
    ×

    苏宁商家

    商家:
    君凤文轩图书专营店
    联系:
    • 商品

    • 服务

    • 物流

    搜索店内商品

    商品分类

    商品参数
    • 作者: [美]德梅萃·P. 博塞克斯(Dimitri P. Bertsekas) 著著| [美]德梅萃·P. 博塞克斯(Dimitri P. Bertsekas) 著编| [美]德梅萃·P. 博塞克斯(Dimitri P. Bertsekas) 著译| [美]德梅萃·P. 博塞克斯(Dimitri P. Bertsekas) 著绘
    • 出版社:清华大学出版社
    • 出版时间:2025-04-01
    • 版次:1
    • 印次:1
    • 印刷时间:2025-04-01
    • 开本:其他
    • ISBN:9787302684718
    • 版权提供:清华大学出版社
  • 作者: [美]德梅萃·P. 博塞克斯(Dimitri P. Bertsekas) 著
  • 著: [美]德梅萃·P. 博塞克斯(Dimitri P. Bertsekas) 著 译
  • 装帧: 平装
  • 印次: 1
  • 定价: 79
  • ISBN: 9787302684718
  • 出版社: 清华大学出版社
  • 开本: 其他
  • 印刷时间: 2025-04-01
  • 语种: 暂无
  • 出版时间: 2025-04-01
  • 页数: 0
  • 外部编号: 庄村51591
  • 版次: 1
  • 成品尺寸: 暂无
  • 1.AlphaZero, Off-Line Training, and On-Line Play
    1.1.Off-Line Training and Policy Iteration
    1.2.On-Line Play and Approximation in Value Space-Truncated Rollout
    1.3.The Lessons of AlphaZero
    1.4.A New Conceptual Framework for Reinforcement Learning
    1.5.Notes and Sources
    2.Deterministic and Stochastic Dynamic Programming
    2.1.Optimal Control Over an Infinite Horizon
    2.2.Approximation in Value Space
    2.3.Notes and Sources
    3.An Abstract View of Reinforcement Learning
    3.1.Bellman Operators
    3.2.Approximation in Value Space and Newton's Method
    3.3.Region of Stability
    3.4.Policy Iteration, Rollout, and Newton's Method
    3.5.How Sensitive is On-Line Play to the Off-Line Training Process?
    3.6.Why Not Just Train a Policy Network and Use it Without On-Line Play?
    3.7.Multiagent Problems and Multiagent Rollout
    3.8.On-Line Simplified Policy Iteration
    3.9.Exceptional Cases
    3.10.Notes and Sources
    4.The Linear Quadratic Case - Illustrations
    4.1.Optimal Solution
    4.2.Cost Functions of Stable Linear Policies
    4.3.Value Iteration
    4.4.One-Step and Multistep Lookahead - Newton Step Interpretations
    4.5.Sensitivity Issues
    4.6.Rollout and Policy Iteration
    4.7.Truncated Rollout - Length of Lookahead Issues
    4.8.Exceptional Behavior in Linear Quadratic Problems
    4.9.Notes and Sources
    5.Adaptive and Model Predictive Control
    5.1.Systems with Unknown Parameters - Robust and PID Control
    5.2.Approximation in Value Space, Rollout, and Adaptive Control
    5.3.Approximation in Value Space, Rollout, and Model Predictive Control
    5.4.Terminal Cost Approximation - Stability Issues
    5.5.Notes and Sources
    6.Finite Horizon Deterministic Problems - Discrete Optimization
    6.1.Deterministic Discrete Spaces Finite Horizon Problems.
    6.2.General Discrete Optimization Problems
    6.3.Approximation in Value Space
    6.4.Rollout Algorithms for Discrete Optimization .. .
    6.5.Rollout and Approximation in Value Space with Multistep Lookahead
    6.5.1.Simplified Multistep Rollout - Double Rollout..p.
    6.5.2.Incremental Rollout for Multistep Approximation in Value Space
    6.6.Constrained Forms of Rollout Algorithms
    6.7.Adaptive Control by Rollout with a POMDP Formulation
    6.8.Rollout for Minimax Control
    6.9.Small Stage Costs and Long Horizon - Continuous-Time Rollout
    6.10.Epilogue
    Appendix A: Newton's Method and Error Bounds
    A.1.Newton's Method for Differentiable Fixed Point Problems
    A.2.Newton's Method Without Differentiability of the Hellman Operator
    A.3.Local and Global Error Bounds for Approximation in Value Space
    A.4.Local and Global Error Bounds for Approximate Policy Iteration
    References

    售后保障

    最近浏览

    猜你喜欢

    该商品在当前城市正在进行 促销

    注:参加抢购将不再享受其他优惠活动

    x
    您已成功将商品加入收藏夹

    查看我的收藏夹

    确定

    非常抱歉,您前期未参加预订活动,
    无法支付尾款哦!

    关闭

    抱歉,您暂无任性付资格

    此时为正式期SUPER会员专享抢购期,普通会员暂不可抢购