• Fri frakt över 249 kr
  • •
  • Snabba leveranser
  • •
  • Billiga böcker
Kundservice

Du är på sajten för privatpersoner.

Företag, bibliotek eller offentlig verksamhet?

Du handlar på classic.bokus.com, där alla dina funktioner finns intakta.
Till classic.bokus.com
Bokus logotyp. Gå till startsidan.
  • Erbjudanden
  • Nyheter
  • Student
  • Topplistor
  • Barn & ungdom
  • Bokus Play
  • E-böcker
  • Pocketböcker
  • Spel & pussel

10% studentrabatt med kod TERM26

Sidfot

Mina sidor

    Hjälp

    • Kundservice
    • Vanliga frågor och svar
    • Frakt och leverans
    • Retur vid ångerrätt
    • Reklamera vara
    • Betalning
    • Köpvillkor
    • Allmänna villkor
    • Information om webbplatsens tillgänglighet

    Om Bokus

    • Om oss
    • Pressrum
    • För studenter
    • För företag
    • För bibliotek och offentlig verksamhet
    • För leverantörer
    • Hållbarhet

    Populärt

    • Aktuella erbjudanden
    • Presentkort
    • Studentlitteratur
    • Nya böcker
    • Topplistor
    • Signerade böcker
    • Engelska böcker

    Inspiration

    • Boktips
    • BookTok
    • Populära bokserier
    • Barnbokskaraktärer
    • Populära författare

    Mina sidor

      Hjälp

      • Kundservice
      • Vanliga frågor och svar
      • Frakt och leverans
      • Retur vid ångerrätt
      • Reklamera vara
      • Betalning
      • Köpvillkor
      • Allmänna villkor
      • Information om webbplatsens tillgänglighet

      Om Bokus

      • Om oss
      • Pressrum
      • För studenter
      • För företag
      • För bibliotek och offentlig verksamhet
      • För leverantörer
      • Hållbarhet

      Populärt

      • Aktuella erbjudanden
      • Presentkort
      • Studentlitteratur
      • Nya böcker
      • Topplistor
      • Signerade böcker
      • Engelska böcker

      Inspiration

      • Boktips
      • BookTok
      • Populära bokserier
      • Barnbokskaraktärer
      • Populära författare
      Logotyp för Bokus
      Följ oss på Facebook (extern länk)Följ oss på Instagram (extern länk)Följ oss på YouTube (extern länk)Följ oss på TikTok (extern länk)
      bokus @ CookiesAnpassa cookiesIntegritetspolicyKöpvillkor
      Till Citymail hemsida (extern länk)Till Budbee hemsida (extern länk)Till Postnord hemsida (extern länk)Till Schenker hemsida (extern länk)Till Early Bird hemsida (extern länk)Till Walleys hemsida (extern länk)
      1. Naturvetenskap och teknik
      2. Matematik och naturvetenskap
      3. Matematik
      4. Tillämpad matematik

      Approximate Dynamic Programming

      Solving the Curses of Dimensionality

      AvWarren B. Powell

      Inbunden, Engelska, 2011

      Del 842 i serien Wiley Series in Probability and Statistics

      1 635 kr

      Beställningsvara. Skickas inom 5-8 vardagar. Fri frakt över 249 kr.

      Fler format och utgåvor

      E-bok

      1 909 kr

      E-bok

      1 877 kr

      Beskrivning

      Praise for the First Edition "Finally, a book devoted to dynamic programming and written using the language of operations research (OR)! This beautiful book fills a gap in the libraries of OR specialists and practitioners."—Computing ReviewsThis new edition showcases a focus on modeling and computation for complex classes of approximate dynamic programming problemsUnderstanding approximate dynamic programming (ADP) is vital in order to develop practical and high-quality solutions to complex industrial problems, particularly when those problems involve making decisions in the presence of uncertainty. Approximate Dynamic Programming, Second Edition uniquely integrates four distinct disciplines—Markov decision processes, mathematical programming, simulation, and statistics—to demonstrate how to successfully approach, model, and solve a wide range of real-life problems using ADP.The book continues to bridge the gap between computer science, simulation, and operations research and now adopts the notation and vocabulary of reinforcement learning as well as stochastic search and simulation optimization. The author outlines the essential algorithms that serve as a starting point in the design of practical solutions for real problems. The three curses of dimensionality that impact complex problems are introduced and detailed coverage of implementation challenges is provided. The Second Edition also features: A new chapter describing four fundamental classes of policies for working with diverse stochastic optimization problems: myopic policies, look-ahead policies, policy function approximations, and policies based on value function approximations A new chapter on policy search that brings together stochastic search and simulation optimization concepts and introduces a new class of optimal learning strategies Updated coverage of the exploration exploitation problem in ADP, now including a recently developed method for doing active learning in the presence of a physical state, using the concept of the knowledge gradient A new sequence of chapters describing statistical methods for approximating value functions, estimating the value of a fixed policy, and value function approximation while searching for optimal policies The presented coverage of ADP emphasizes models and algorithms, focusing on related applications and computation while also discussing the theoretical side of the topic that explores proofs of convergence and rate of convergence. A related website features an ongoing discussion of the evolving fields of approximation dynamic programming and reinforcement learning, along with additional readings, software, and datasets.Requiring only a basic understanding of statistics and probability, Approximate Dynamic Programming, Second Edition is an excellent book for industrial engineering and operations research courses at the upper-undergraduate and graduate levels. It also serves as a valuable reference for researchers and professionals who utilize dynamic programming, stochastic programming, and control theory to solve problems in their everyday work.

      Produktinformation

      • Utgivningsdatum:2011-11-18
      • Mått:155 x 236 x 41 mm
      • Vikt:1 043 g
      • Format:Inbunden
      • Språk:Engelska
      • Serie:Wiley Series in Probability and Statistics
      • Antal sidor:656
      • Upplaga:2
      • Förlag:John Wiley & Sons Inc
      • ISBN:9780470604458

      Utforska kategorier

      • Tillämpad matematik inom Naturvetenskap och teknik

      Mer om författaren

      WARREN B. POWELL, PhD, is Professor of Operations Research and Financial Engineering at Princeton University, where he is founder and Director of CASTLE Laboratory, a research unit that works with industrial partners to test new ideas found in operations research. The recipient of the 2004 INFORMS Fellow Award, Dr. Powell has authored more than 160 published articles on stochastic optimization, approximate dynamicprogramming, and dynamic resource management.

      Innehållsförteckning

      • Preface to the Second Edition xi Preface to the First Edition xvAcknowledgments xvii1 The Challenges of Dynamic Programming 11.1 A Dynamic Programming Example: A Shortest Path Problem, 21.2 The Three Curses of Dimensionality, 31.3 Some Real Applications, 61.4 Problem Classes, 111.5 The Many Dialects of Dynamic Programming, 151.6 What Is New in This Book?, 171.7 Pedagogy, 191.8 Bibliographic Notes, 222 Some Illustrative Models 252.1 Deterministic Problems, 262.2 Stochastic Problems, 312.3 Information Acquisition Problems, 472.4 A Simple Modeling Framework for Dynamic Programs, 502.5 Bibliographic Notes, 54Problems, 543 Introduction to Markov Decision Processes 573.1 The Optimality Equations, 583.2 Finite Horizon Problems, 653.3 Infinite Horizon Problems, 663.4 Value Iteration, 683.5 Policy Iteration, 743.6 Hybrid Value-Policy Iteration, 753.7 Average Reward Dynamic Programming, 763.8 The Linear Programming Method for Dynamic Programs, 773.9 Monotone Policies*, 783.10 Why Does It Work?**, 843.11 Bibliographic Notes, 103Problems, 1034 Introduction to Approximate Dynamic Programming 1114.1 The Three Curses of Dimensionality (Revisited), 1124.2 The Basic Idea, 1144.3 Q-Learning and SARSA, 1224.4 Real-Time Dynamic Programming, 1264.5 Approximate Value Iteration, 1274.6 The Post-Decision State Variable, 1294.7 Low-Dimensional Representations of Value Functions, 1444.8 So Just What Is Approximate Dynamic Programming?, 1464.9 Experimental Issues, 1494.10 But Does It Work?, 1554.11 Bibliographic Notes, 156Problems, 1585 Modeling Dynamic Programs 1675.1 Notational Style, 1695.2 Modeling Time, 1705.3 Modeling Resources, 1745.4 The States of Our System, 1785.5 Modeling Decisions, 1875.6 The Exogenous Information Process, 1895.7 The Transition Function, 1985.8 The Objective Function, 2065.9 A Measure-Theoretic View of Information**, 2115.10 Bibliographic Notes, 213Problems, 2146 Policies 2216.1 Myopic Policies, 2246.2 Lookahead Policies, 2246.3 Policy Function Approximations, 2326.4 Value Function Approximations, 2356.5 Hybrid Strategies, 2396.6 Randomized Policies, 2426.7 How to Choose a Policy?, 2446.8 Bibliographic Notes, 247Problems, 2477 Policy Search 2497.1 Background, 2507.2 Gradient Search, 2537.3 Direct Policy Search for Finite Alternatives, 2567.4 The Knowledge Gradient Algorithm for Discrete Alternatives, 2627.5 Simulation Optimization, 2707.6 Why Does It Work?**, 2747.7 Bibliographic Notes, 285Problems, 2868 Approximating Value Functions 2898.1 Lookup Tables and Aggregation, 2908.2 Parametric Models, 3048.3 Regression Variations, 3148.4 Nonparametric Models, 3168.5 Approximations and the Curse of Dimensionality, 3258.6 Why Does It Work?**, 3288.7 Bibliographic Notes, 333Problems, 3349 Learning Value Function Approximations 3379.1 Sampling the Value of a Policy, 3379.2 Stochastic Approximation Methods, 3479.3 Recursive Least Squares for Linear Models, 3499.4 Temporal Difference Learning with a Linear Model, 3569.5 Bellman’s Equation Using a Linear Model, 3589.6 Analysis of TD(0), LSTD, and LSPE Using a Single State, 3649.7 Gradient-Based Methods for Approximate Value Iteration*, 3669.8 Least Squares Temporal Differencing with Kernel Regression*, 3719.9 Value Function Approximations Based on Bayesian Learning*, 3739.10 Why Does It Work*, 3769.11 Bibliographic Notes, 379Problems, 38110 Optimizing While Learning 38310.1 Overview of Algorithmic Strategies, 38510.2 Approximate Value Iteration and Q-Learning Using Lookup Tables, 38610.3 Statistical Bias in the Max Operator, 39710.4 Approximate Value Iteration and Q-Learning Using Linear Models, 40010.5 Approximate Policy Iteration, 40210.6 The Actor–Critic Paradigm, 40810.7 Policy Gradient Methods, 41010.8 The Linear Programming Method Using Basis Functions, 41110.9 Approximate Policy Iteration Using Kernel Regression*, 41310.10 Finite Horizon Approximations for Steady-State Applications, 41510.11 Bibliographic Notes, 416Problems, 41811 Adaptive Estimation and Stepsizes 41911.1 Learning Algorithms and Stepsizes, 42011.2 Deterministic Stepsize Recipes, 42511.3 Stochastic Stepsizes, 43311.4 Optimal Stepsizes for Nonstationary Time Series, 43711.5 Optimal Stepsizes for Approximate Value Iteration, 44711.6 Convergence, 44911.7 Guidelines for Choosing Stepsize Formulas, 45111.8 Bibliographic Notes, 452Problems, 45312 Exploration Versus Exploitation 45712.1 A Learning Exercise: The Nomadic Trucker, 45712.2 An Introduction to Learning, 46012.3 Heuristic Learning Policies, 46412.4 Gittins Indexes for Online Learning, 47012.5 The Knowledge Gradient Policy, 47712.6 Learning with a Physical State, 48212.7 Bibliographic Notes, 492Problems, 49313 Value Function Approximations for Resource Allocation Problems 49713.1 Value Functions versus Gradients, 49813.2 Linear Approximations, 49913.3 Piecewise-Linear Approximations, 50113.4 Solving a Resource Allocation Problem Using Piecewise-Linear Functions, 50513.5 The SHAPE Algorithm, 50913.6 Regression Methods, 51313.7 Cutting Planes*, 51613.8 Why Does It Work?**, 52813.9 Bibliographic Notes, 535Problems, 53614 Dynamic Resource Allocation Problems 54114.1 An Asset Acquisition Problem, 54114.2 The Blood Management Problem, 54714.3 A Portfolio Optimization Problem, 55714.4 A General Resource Allocation Problem, 56014.5 A Fleet Management Problem, 57314.6 A Driver Management Problem, 58014.7 Bibliographic Notes, 585Problems, 58615 Implementation Challenges 59315.1 Will ADP Work for Your Problem?, 59315.2 Designing an ADP Algorithm for Complex Problems, 59415.3 Debugging an ADP Algorithm, 59615.4 Practical Issues, 59715.5 Modeling Your Problem, 60215.6 Online versus Offline Models, 60415.7 If It Works, Patent It!, 606Bibliography 607Index 623
      Hoppa över listan

      Mer från samma författare

      Warren B. Powell, Ilya O. Ryzhov - Optimal Learning, Inbunden
      Del 841

      Optimal Learning

      Warren B. Powell, Ilya O. Ryzhov

      Inbunden, 2012

      1 432 kr

      Jennie Si, Andrew G. Barto, Warren B. Powell, Don Wunsch - Handbook of Learning and Approximate Dynamic Programming, Inbunden
      Del 2

      Handbook of Learning and Approximate Dynamic Programming

      Jennie Si, Andrew G. Barto, Warren B. Powell, Don Wunsch

      Inbunden, 2004

      2 130 kr

      Ilya O. Ryzhov, Warren B. Powell - Optimal Learning, E-bok

      Optimal Learning

      Ilya O. Ryzhov, Warren B. Powell

      E-bok
      2012

      1 656 kr

      Ilya O. Ryzhov, Warren B. Powell - Optimal Learning, E-bok

      Optimal Learning

      Ilya O. Ryzhov, Warren B. Powell

      E-bok
      2013

      1 656 kr

      Warren B. Powell - Reinforcement Learning and Stochastic Optimization, Inbunden

      Reinforcement Learning and Stochastic Optimization

      Warren B. Powell

      Inbunden, 2022

      1 698 kr

      Warren B. Powell - Reinforcement Learning and Stochastic Optimization, E-bok

      Reinforcement Learning and Stochastic Optimization

      Warren B. Powell

      E-bok
      2022

      1 909 kr

      Warren B. Powell - Reinforcement Learning and Stochastic Optimization, E-bok

      Reinforcement Learning and Stochastic Optimization

      Warren B. Powell

      E-bok
      2022

      1 898 kr

      Warren B. Powell - Sequential Decision Analytics and Modeling, Inbunden
      Del 42

      Sequential Decision Analytics and Modeling

      Warren B. Powell

      Inbunden, 2022

      1 038 kr

      Hoppa över listan

      Mer från samma serie

      Theodore W. Anderson - Introduction to Multivariate Statistical Analysis, Inbunden
      • -10% student
      Del 355

      Introduction to Multivariate Statistical Analysis

      Theodore W. Anderson

      Inbunden, 2003

      2 305 kr

      José M. Bernardo, Adrian F. M. Smith - Bayesian Theory, Inbunden
      • -10% student
      Del 316

      Bayesian Theory

      José M. Bernardo, Adrian F. M. Smith

      Inbunden, 1994

      4 846 kr

      David W. Scott - Multivariate Density Estimation, Inbunden
      • -10% student

      Multivariate Density Estimation

      David W. Scott

      Inbunden, 2015

      1 310 kr

      Saltelli, Chan, Andrea Saltelli, K. Chan, E. M. Scott - Sensitivity Analysis, Inbunden
      Del 535

      Sensitivity Analysis

      Saltelli, Chan, Andrea Saltelli, K. Chan, E. M. Scott

      Inbunden, 2000

      2 057 kr

      Walter Enders - Applied Econometric Time Series, Häftad

      Applied Econometric Time Series

      Walter Enders

      Häftad, 2014

      2 995 kr

      Sanjeev Kulkarni, Gilbert Harman - Elementary Introduction to Statistical Learning Theory, Inbunden
      Del 853

      Elementary Introduction to Statistical Learning Theory

      Sanjeev Kulkarni, Gilbert Harman

      Inbunden, 2011

      1 375 kr

      John W. Lamperti - Probability, Inbunden
      Del 310

      Probability

      John W. Lamperti

      Inbunden, 1996

      2 142 kr

      George E. P. Box, Gwilym M. Jenkins, Gregory C. Reinsel, Greta M. Ljung - Time Series Analysis, Inbunden

      Time Series Analysis

      George E. P. Box, Gwilym M. Jenkins, Gregory C. Reinsel, Greta M. Ljung

      Inbunden, 2015

      5,0 utav 5 stjärnor. Totalt antal röster:(1)

      1 759 kr

      Ali S. Hadi, Samprit Chatterjee - Regression Analysis By Example Using R, Inbunden
      Del 99

      Regression Analysis By Example Using R

      Ali S. Hadi, Samprit Chatterjee

      Inbunden, 2023

      1 544 kr

      Douglas C. Montgomery, Cheryl L. Jennings, Murat Kulahci - Introduction to Time Series Analysis and Forecasting, 1e Student Solutions Manual, Häftad
      Del 763

      Introduction to Time Series Analysis and Forecasting, 1e Student Solutions Manual

      Douglas C. Montgomery, Cheryl L. Jennings, Murat Kulahci

      Häftad, 2009

      513 kr

      Hoppa över listan

      Du kanske också är intresserad av

      Piet Vanassche, Georges Gielen, Willy M Sansen - Systematic Modeling and Analysis of Telecom Frontends and their Building Blocks, Häftad
      Del 842

      Systematic Modeling and Analysis of Telecom Frontends and their Building Blocks

      Piet Vanassche, Georges Gielen, Willy M Sansen

      Häftad, 2011

      1 634 kr

      Oliver Green - Trams and Trolleybuses, Häftad
      Del 842

      Trams and Trolleybuses

      Oliver Green

      Häftad, 2018

      112 kr

      José Bravo, Gabriel Urzáiz - Proceedings of the 15th International Conference on Ubiquitous Computing & Ambient Intelligence (UCAmI 2023), Häftad
      Del 842

      Proceedings of the 15th International Conference on Ubiquitous Computing & Ambient Intelligence (UCAmI 2023)

      José Bravo, Gabriel Urzáiz

      Häftad, 2023

      2 242 kr

      Waubgeshig Rice - Mond des verharschten Schnees, Häftad
      Del 842

      Mond des verharschten Schnees

      Waubgeshig Rice

      Häftad, 2024

      184 kr

      Bernd Holla - Qualitaetsentwicklung in Der Weiterbildung Durch Praxisorientierte Evaluation, Häftad
      Del 842

      Qualitaetsentwicklung in Der Weiterbildung Durch Praxisorientierte Evaluation

      Bernd Holla

      Häftad, 2002

      709 kr

      Carl Kröger - Motorabgase und ihre Reinigung, Häftad
      Del 842

      Motorabgase und ihre Reinigung

      Carl Kröger

      Häftad, 1960

      565 kr

      Oliver Green - Trams and Trolleybuses, E-bok
      Del 842

      Trams and Trolleybuses

      Oliver Green

      E-bok
      2018

      106 kr

      Jemal Abawajy, Kim-Kwang Raymond Choo, Rafiqul Islam, Zheng Xu, Mohammed Atiquzzaman - International Conference on Applications and Techniques in Cyber Security and Intelligence ATCI 2018, Övrigt
      Del 842

      International Conference on Applications and Techniques in Cyber Security and Intelligence ATCI 2018

      Jemal Abawajy, Kim-Kwang Raymond Choo, Rafiqul Islam, Zheng Xu, Mohammed Atiquzzaman

      2 643 kr

      Sauerland 2 + Aktiv Guide, Övrigt
      Del 842

      Sauerland 2 + Aktiv Guide

      Kompass Karten GmbH

      161 kr

      Zainah Md. Zain, Mohd. Herwan Sulaiman, Amir Izzani Mohamed, Mohd. Shafie Bakar, Mohd. Syakirin Ramli - Proceedings of the 6th International Conference on Electrical, Control and Computer Engineering, Övrigt
      Del 842

      Proceedings of the 6th International Conference on Electrical, Control and Computer Engineering

      Zainah Md. Zain, Mohd. Herwan Sulaiman, Amir Izzani Mohamed, Mohd. Shafie Bakar, Mohd. Syakirin Ramli

      2 119 kr