A Multi-Modal Story Generation Framework with AI-Driven Storyline Guidance

Citations

WEB OF SCIENCE

7
Citations

SCOPUS

11

초록

An automatic story generation system continuously generates stories with a natural plot. The major challenge of automatic story generation is to maintain coherence between consecutive generated stories without the need for human intervention. To address this, we propose a novel multi-modal story generation framework that includes automated storyline decision-making capabilities. Our framework consists of three independent models: a transformer encoder-based storyline guidance model, which predicts a storyline using a multiple-choice question-answering problem; a transformer decoder-based story generation model that creates a story that describes the storyline determined by the guidance model; and a diffusion-based story visualization model that generates a representative image visually describing a scene to help readers better understand the story flow. Our proposed framework was extensively evaluated through both automatic and human evaluations, which demonstrate that our model outperforms the previous approach, suggesting the effectiveness of our storyline guidance model in making proper plans.

키워드

multi-modal story generationAI-driven storyline guidancemultiple-choice question answeringautomatic story generationstory visualizationdiffusion
제목
A Multi-Modal Story Generation Framework with AI-Driven Storyline Guidance
저자
Kim, JuntaeHeo, YoonseokYu, HogeonNang, Jongho
DOI
10.3390/electronics12061289
발행일
2023-03
유형
Article
저널명
Electronics (Basel)
12
6