Using Large Language Models for Expert Prior Elicitation in Predictive Modelling

stp2yDecember 11, 20240 Comments

AmazUtah_NLP at SemEval-2024 Task 9: A MultiChoice Question Answering System for Commonsense Defying Reasoning

[Submitted on 26 Nov 2024 (v1), last revised 10 Dec 2024 (this version, v2)]

View a PDF of the paper titled Using Large Language Models for Expert Prior Elicitation in Predictive Modelling, by Alexander Capstick and 2 other authors

View PDF
HTML (experimental)

Abstract:Large language models (LLMs), trained on diverse data effectively acquire a breadth of information across various domains. However, their computational complexity, cost, and lack of transparency hinder their direct application for specialised tasks. In fields such as clinical research, acquiring expert annotations or prior knowledge about predictive models is often costly and time-consuming. This study proposes the use of LLMs to elicit expert prior distributions for predictive models. This approach also provides an alternative to in-context learning, where language models are tasked with making predictions directly. In this work, we compare LLM-elicited and uninformative priors, evaluate whether LLMs truthfully generate parameter distributions, and propose a model selection strategy for in-context learning and prior elicitation. Our findings show that LLM-elicited prior parameter distributions significantly reduce predictive error compared to uninformative priors in low-data settings. Applied to clinical problems, this translates to fewer required biological samples, lowering cost and resources. Prior elicitation also consistently outperforms and proves more reliable than in-context learning at a lower cost, making it a preferred alternative in our setting. We demonstrate the utility of this method across various use cases, including clinical applications. For infection prediction, using LLM-elicited priors reduced the number of required labels to achieve the same accuracy as an uninformative prior by 55%, 200 days earlier in the study.

Submission history

From: Alexander Capstick [view email]
[v1]
Tue, 26 Nov 2024 10:13:39 UTC (3,069 KB)
[v2]
Tue, 10 Dec 2024 11:36:48 UTC (3,069 KB)

Source link
lol

By stp2y