Dream Scene Visualiser(DSV)는 글로 쓴 꿈 묘사를 4개 패널의 시간순 이미지 시퀀스로 시각화하는 시스템이다. 먼저 LLM이 꿈 묘사를 4개의 시간순 부분으로 분할하고, 텍스트-이미지 모델이 시퀀스 전체의 시각적 일관성을 유지하며 각 부분의 이미지를 생성하며, 텍스트와 맞지 않는 이미지는 DSV가 재생성한다. DreamBank의 꿈 묘사 50건으로 시각화를 생성해 CLIP, DINOv2, Qwen2-VL 비전-언어 모델을 활용한 객관적 지표로 품질·충실도·일관성을 평가했다. 표현하기 어려운 감정적으로 강렬한 꿈 경험을 구조화된 이미지 시퀀스로 전환해 소통 가능하게 만든다는 점에서 의의가 있다.
- •LLM으로 꿈 묘사를 4개 시간순 부분으로 자동 분할
- •텍스트-이미지 모델로 시퀀스 전체의 시각적 일관성을 유지하며 이미지 생성
- •텍스트와 불일치하는 이미지는 자동 재생성
- •DreamBank 50건 시각화를 CLIP·DINOv2·Qwen2-VL로 정량 평가
- •ICCC 2026 단편 논문으로 채택
0단 자동
AI가 규칙대로 쓰고 그대로 게시했습니다. 사람이 따로 보지 않았습니다.
- 규칙 판
- 규칙 판 도입 이전 기사입니다.
- 남기는 것
- 규칙 판 · 모델 · 시각
- 판 기록
- 아직 없습니다.
Coherence-Oriented Dream Scene Visualisation
본문 미리보기
arXiv:2608.05233v1 Announce Type: new Abstract: Dreams can be emotionally intense but difficult to communicate. We describe the Dream Scene Visualiser (DSV) system which turns written dream descriptions into a temporal sequence of four panel images visualising the dream. This starts with a large language model prompted to split a dream description into four chronological parts. Then a text-to-image model produces images for each part with visual coherence maintained across the sequence, and DSV
전체 내용이 궁금하다면?
원문을 직접 읽어보세요
이 글이 만들어진 과정
- 10:11AI 초안

