Human and Machine Learning in Non-Markovian Decision Making

Clarke, Aaron Michael; Friedrich, Johannes; Tartaglia, Elisa M.; Marchesotti, Silvia; Senn, Walter; Herzog, Michael H.

Human and Machine Learning in Non-Markovian Decision Making

Clarke, Aaron Michael; Friedrich, Johannes; Tartaglia, Elisa M.; Marchesotti, Silvia; Senn, Walter; Herzog, Michael H.

2015

Formats

Format
BibTeX
MARCXML
TextMARC
MARC
DataCite
DublinCore
EndNote
NLM
RefWorks
RIS

Add to Basket

Cite

Files

Abstract

Humans can learn under a wide variety of feedback conditions. Reinforcement learning (RL), where a series of rewarded decisions must be made, is a particularly important type of learning. Computational and behavioral studies of RL have focused mainly on Markovian decision processes, where the next state depends on only the current state and action. Little is known about non-Markovian decision making, where the next state depends on more than the current state and action. Learning is non-Markovian, for example, when there is no unique mapping between actions and feedback. We have produced a model based on spiking neurons that can handle these non-Markovian conditions by performing policy gradient descent [1]. Here, we examine the model’s performance and compare it with human learning and a Bayes optimal reference, which provides an upper-bound on performance. We find that in all cases, our population of spiking neurons model well-describes human performance.

Details

Title

Human and Machine Learning in Non-Markovian Decision Making

Author

Clarke, Aaron Michael : École Polytechnique Fédérale de Lausanne
Friedrich, Johannes : University of Berne
Tartaglia, Elisa M. : University of Chicago
Marchesotti, Silvia : École Polytechnique Fédérale de Lausanne
Senn, Walter : University of Berne
Herzog, Michael H. : École Polytechnique Fédérale de Lausanne

Content Type

Article

Published in

PLOS ONE

Identifier(s)

DOI: https://doi.org/10.1371/journal.pone.0123105

Data availability statement

All relevant data are plotted in the manuscript. In addition, the raw data are also available at the Open Science Framework (https://osf.io/login/?next=/9sacv/): https://osf.io/9sacv/?view_only=187a993a964342cfab1e8f7f65678fa9.

Funding Information

Swiss National Science Foundation, Learning from delayed and sparse feedback
Human Brain Project
SystemsX.ch
Swiss National Science Foundation, Perspective Researcher fellowship, PBBEP3-146112
ProDoc, Top-down and bottom-up processes in perceptual learning, PDFM11-114404
Swiss National Science Foundation, Perspective Researcher fellowship, PBELP3-135838

Publication Date

2015-04-21

Language

English

Copyright Statement

This is an open access article distributed under the terms of the Creative Commons Attribution License, which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are credited

Licensing

CC BY

Record Appears in

Biological Sciences Division > Neurobiology
Physical Sciences Division > Statistics
All

Record Created

2024-01-09

Preview

Statistics

Download Full History