Human and Machine Learning in Non-Markovian Decision Making

Humans can learn under a wide variety of feedback conditions. Reinforcement learning (RL), where a series of rewarded decisions must be made, is a particularly important type of learning. Computational and behavioral studies of RL have focused mainly on Markovian decision processes, where the next state depends on only the current state and action. Little is known about non-Markovian decision making, where the next state depends on more than the current state and action. Learning is non-Markovian, for example, when there is no unique mapping between actions and feedback. We have produced a model based on spiking neurons that can handle these non-Markovian conditions by performing policy gradient descent [1]. Here, we examine the model's performance and compare it with human learning and a Bayes optimal reference, which provides an upper-bound on performance. We find that in all cases, our population of spiking neurons model well-describes human performance.

Chattez avec Graph Search

Posez n’importe quelle question sur les cours, conférences, exercices, recherches, actualités, etc. de l’EPFL ou essayez les exemples de questions ci-dessous.

AVERTISSEMENT : Le chatbot Graph n'est pas programmé pour fournir des réponses explicites ou catégoriques à vos questions. Il transforme plutôt vos questions en demandes API qui sont distribuées aux différents services informatiques officiellement administrés par l'EPFL. Son but est uniquement de collecter et de recommander des références pertinentes à des contenus que vous pouvez explorer pour vous aider à répondre à vos questions.

Human and Machine Learning in Non-Markovian Decision Making

Graph Chatbot

Chattez avec Graph Search

Seeking the new, learning from the unexpected: Computational models of surprise and novelty in the brain

From event-based surprise to lifelong learning.A journey in the timescales of adaptation

Breaking the Curse of Dimensionality in Deep Neural Networks by Learning Invariant Representations

Seeking the new, learning from the unexpected: Computational models of surprise and novelty in the brain

Breaking the Curse of Dimensionality in Deep Neural Networks by Learning Invariant Representations

From event-based surprise to lifelong learning.A journey in the timescales of adaptation