Temporal-difference methods and Markov models

E Barnard

doi:10.1109/21.229449

Temporal-difference methods and Markov models

E Barnard

https://doi.org/10.1109/21.229449

Copy DOI

Journal: IEEE Transactions on Systems, Man, and Cybernetics	Publication Date: Jan 1, 1993
Citations: 44

Affiliation: University of Pretoria

#Temporal-difference Methods #Markov Models + Show 7 more

Abstract
Full-Text PDF
Similar Papers

Abstract

The relation between temporal-difference training methods and Markov models is explored. This relation is derived from a new perspective, and in this way the particular association between conventional temporal-difference methods and first-order Markov models is explained. The authors then derive a generalization of temporal-difference methods that is suitable for Markov models of higher order. Several issues related to the performance of mismatched temporal-difference methods (i.e., the performance when the temporal-difference method is not specifically designed to match the order of the Markov model) are investigated numerically. >

Full Text

Paper version not known

Open DOI Link

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Similar Papers

Paper Title

Journal

Date

Author

View more papers

More From: IEEE Transactions on Systems, Man, and Cybernetics

Paper Title

Journal

Date

Author

View more papers

Disclaimer: All third-party content on this website/platform is and will remain the property of their respective owners and is provided on "as is" basis without any warranties, express or implied. Use of third-party content does not indicate any affiliation, sponsorship with or endorsement by them. Any references to third-party content is to identify the corresponding services and shall be considered fair use under The CopyrightLaw.