Using large language models to create narrative events

被引：0

作者：

Bartalesi, Valentina ^{[1
]}

Lenzi, Emanuele ^{[1
,2
]}

De Martino, Claudio ^{[1
]}

机构：

[1] Natl Res Council Italy CNR, Inst Informat Sci & Technol Alessandro Faedo ISTI, Pisa, Italy

[2] Univ Pisa, Dept Informat Engn DII, Pisa, Italy

来源：

PEERJ COMPUTER SCIENCE | 2024年 / 10卷

关键词：

Large language models; Narratives; Events; Semantic web; Digital humanities; SCIENCE;

D O I：

10.7717/peerj-cs.2242

中图分类号：

TP18 [人工智能理论];

学科分类号：

081104 ; 0812 ; 0835 ; 1405 ;

摘要：

Narratives play a crucial role in human communication, serving as a means to convey experiences, perspectives, and meanings across various domains. They are particularly significant in scientific communities, where narratives are often utilized to explain complex phenomena and share knowledge. This article explores the possibility of integrating large language models (LLMs) into a workflow that, exploiting the Semantic Web technologies, transforms raw textual data gathered by scientific communities into narratives. In particular, we focus on using LLMs to automatically create narrative events, maintaining the reliability of the generated texts. The study provides a conceptual definition of narrative events and evaluates the performance of different smaller LLMs compared to the requirements we identified. A key aspect of the experiment is the emphasis on maintaining the integrity of the original narratives in the LLM outputs, as experts often review texts produced by scientific communities to ensure their accuracy and reliability. We first perform an evaluation on a corpus of five narratives and then on a larger dataset comprising 124 narratives. LLaMA 2 is identified as the most suitable model for generating narrative events that closely align with the input texts, demonstrating its ability to generate high-quality narrative events. Prompt engineering techniques are then employed to enhance the performance of the selected model, leading to further improvements in the quality of the generated texts.

引用

页数：17

共 42 条

[31]

OpenAI, 2024, Prompt Engineering

[32] Literary genres as norms and good habits [J].

Pavel, T .

NEW LITERARY HISTORY, 2003, 34 (02) :201-210

[33]

Propp Valdimir., 1968, MORPHOLOGY FOLKTALE, V2nd

[34]

Ramlochan S, 2024, System prompts in large language models

[35] The probabilistic basis of Jaccard's index of similarity [J].

Real, R ;

Vargas, JM .

SYSTEMATIC BIOLOGY, 1996, 45 (03) :380-385

[36]

Riccio D., 2023, Extending context length in large language models

[37]

Schiff B, 2012, NARRAT WORKS, V2, P33

[38]

Shklovsky Victor., 2017, LIT THEORY ANTHOLOGY, P8

[39]

Taylor Charles., 1989, Sources of the Self

[40]

Walker C, 2005, Automatic content extraction (ACE) 2005 Multilingual Training Corpus LDC2006T06, DOI [10.35111/mwxc-vh88, DOI 10.35111/MWXC-VH88]

← 1 2 3 4 5 →