StoryER: Automatic Story Evaluation via Ranking, Rating and Reasoning

Chen, Hong; Vo, Duc Minh; Takamura, Hiroya; Miyao, Yusuke; Nakayama, Hideki

Abstract:Existing automatic story evaluation methods place a premium on story lexical level coherence, deviating from human preference. We go beyond this limitation by considering a novel \textbf{Story} \textbf{E}valuation method that mimics human preference when judging a story, namely \textbf{StoryER}, which consists of three sub-tasks: \textbf{R}anking, \textbf{R}ating and \textbf{R}easoning. Given either a machine-generated or a human-written story, StoryER requires the machine to output 1) a preference score that corresponds to human preference, 2) specific ratings and their corresponding confidences and 3) comments for various aspects (e.g., opening, character-shaping). To support these tasks, we introduce a well-annotated dataset comprising (i) 100k ranked story pairs; and (ii) a set of 46k ratings and comments on various aspects of the story. We finetune Longformer-Encoder-Decoder (LED) on the collected dataset, with the encoder responsible for preference score and aspect prediction and the decoder for comment generation. Our comprehensive experiments result in a competitive benchmark for each task, showing the high correlation to human preference. In addition, we have witnessed the joint learning of the preference scores, the aspect ratings, and the comments brings gain in each single task. Our dataset and benchmarks are publicly available to advance the research of story evaluation tasks.\footnote{Dataset and pre-trained model demo are available at anonymous website \url{this http URL} and \url{this https URL}}

Comments:	accepted by EMNLP 2022
Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2210.08459 [cs.CL]
	(or arXiv:2210.08459v2 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2210.08459

Computer Science > Computation and Language

Title:StoryER: Automatic Story Evaluation via Ranking, Rating and Reasoning

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators