Back to results

Massachusetts Institute of Technology

Observational learning with finite memory

Abstract

dc:description.abstract

We study a model of sequential decision making under uncertainty by a population of agents. Each agent prior to making a decision receives a private signal regarding a binary underlying state of the world. Moreover she observes the actions of her last K immediate predecessors. We discriminate between the cases of bounded and unbounded informativeness of private signals. In contrast to the literature that typically assumes myopic agents who choose the action that maximizes the probability of making the correct decision (the decision that identifies correctly the underlying state), in our model we assume that agents are forward looking, maximizing the discounted sum of the probabilities of a correct decision from all the future agents including theirs. Therefore, an agent when making a decision takes into account the impact that this decision will have on the subsequent agents. We investigate whether in a Perfect Bayesian Equilibrium of this model individual's decisions converge to the correct state of the world, in probability, and we show that this cannot happen for any K and any discount factor if private signals' informativeness is bounded. As a benchmark, we analyze the design limits associated with this problem, which entail constructing decision profiles that dictate each agent's action as a function of her information set, given by her private signal and the last K decisions. We investigate the case of bounded informativeness of the private signals. We answer the question whether there exists a decision profile that results in agents' actions converging to the correct state of the world, a property that we call learning. We first study almost sure learning and prove that it is impossible under any decision rule. We then explore learning in probability, where a dichotomy arises. Specifically, if K = 1 we show that learning in probability is impossible under any decision rule, while for K > 2 we design a decision rule that achieves it.

Degree

thesis:*
Department dc:contributor.department
Massachusetts Institute of Technology. Dept. of Electrical Engineering and Computer Science.
Grantor dc:publisher
Massachusetts Institute of Technology
Year dc:date.issued
2011

Author and committee

dc:creator, dc:contributor.*
Author dc:creator
  • Drakopoulos, Kimon
Advisor dc:contributor.advisor
  • Asuman Ozdaglar and John Tsitsiklis.

Subjects

dc:subject × 1

Rights

dc:rights
Statement dc:rights
  • M.I.T. theses are protected by copyright. They may be viewed from this source for any purpose, but reproduction or distribution in any format is prohibited without written permission. See provided URL for inquiries about permission.
Language dc:language.iso
eng

Identifiers

dc:identifier.*
Handle dc:identifier.uri
http://hdl.handle.net/1721.1/66027
OAI identifier oai:identifier
oai:dspace.mit.edu:1721.1/66027

Chain of custody

source
Harvested from
MIT
Base URL
dspace.mit.edu/oai/request
Last updated
2026-07-22
Source record
OAI-PMH GetRecord
citation

Drakopoulos, Kimon. Observational learning with finite memory. Massachusetts Institute of Technology, 2011. http://hdl.handle.net/1721.1/66027