← 피드로
[Submitted on 5 Aug 2026]
Abstract:Dreams can be emotionally intense but difficult to communicate. We describe the Dream Scene Visualiser (DSV) system which turns written dream descriptions into a temporal sequence of four panel images visualising the dream. This starts with a large language model prompted to split a dream description into four chronological parts. Then a text-to-image model produces images for each part with visual coherence maintained across the sequence, and DSV regenerates any image not suitably matching the text. We evaluate DSV over 50 visualisations from dream descriptions in DreamBank, and report quality, fidelity and coherence results via objective measures employing the CLIP, DINOv2 and Qwen2-VL vision-language models.
Submission history
From: Azra Acil [view email]
[v1]
Wed, 5 Aug 2026 13:36:34 UTC (8,812 KB)
추출 본문 · 출처: arxiv.org · https://arxiv.org/abs/2608.05233
답글 남기기