Back to results

University of Cambridge

The Role of Precedent in Computational Models of Law

Abstract

dc:description.abstract

In common law countries, due to the doctrine of precedent, lawyers need to consider a body of case law which grows with each court decision. As a result, the complexity and number of documents a lawyer should be familiar with for each new decision increases rapidly over time. Robust models of precedential reasoning could help in mitigating this problem. Therefore, the goal of this thesis is to model precedent, understand how it works and how to use it to explain system predictions. Towards this, I will use modern techniques from natural language processing (NLP) and machine learning (ML). In my first investigation, I demonstrate how information-theoretic methods can be used to settle a long-standing jurisprudential debate, namely whether it is facts or legal arguments that have a stronger influence on case outcomes. In terms of mutual information, obtained by training a classification neural models built on top of a pre-trained language model, I find that arguments (0.38 nats) have a stronger bearing on outcomes than facts (0.18 nats). I next design a neural model of legal outcome prediction that significantly outperforms prior work. More importantly, it is also able to model negative outcomes: it can learn which law has been claimed as violated but was found not violated, as opposed to only predicting which law has been violated. Negative outcome is important because it forms legally binding precedent, and models that ignore it will inevitably model law incorrectly. My last content chapter is concerned with explainability. Rather than explaining which linguistic faculties a neural model might have learned or how individual words affect the model's behaviour, my operationalisation of this task is more informative from a legal perspective. I use influence functions, which allow me to measure the change in loss to efficiently estimate the effect of removing a single precedent from the model. The results demonstrate that existing outcome-prediction models lack the ability to employ most of the basic types of precedential reasoning. Since there are many outstanding questions about ethics in legal NLP, I elaborate on these in the final chapter of this thesis. In summary, my thesis can be seen as an exercise in applying insights from one discipline to another: modern NLP techniques can advance our understanding of the law, whereas ideas from the legal domain can be applied to build stronger computational models of law.

Degree

thesis:*
Name dc:type.qualificationname
Doctor of Philosophy (PhD)
Level dc:type.qualificationlevel
Doctoral
Grantor dc:publisher.institution
University of Cambridge
Year dc:date.issued
2023

Author and committee

dc:creator, dc:contributor.*
Author dc:creator
  • Valvoda, Josef
Advisor dc:contributor.advisor
  • Teufel, Simone

Subjects

dc:subject × 3

Rights

dc:rights
Language dc:language
eng

Identifiers

dc:identifier.*
DOI dc:identifier.doi
https://doi.org/10.17863/CAM.114276
OAI identifier oai:identifier
oai:www.repository.cam.ac.uk:1810/377462

Chain of custody

source
Harvested from
Cambridge University
Base URL
api.repository.cam.ac.uk/server/oai/request
Last updated
2026-07-22
Source record
OAI-PMH GetRecord
citation

Valvoda, Josef. The Role of Precedent in Computational Models of Law. Doctoral thesis, University of Cambridge, 2023. https://doi.org/10.17863/CAM.114276