Human-Aligned Procedural Level Generation Reinforcement Learning via Text-Level-Sketch Shared Representation

작성자

카테고리:

← 피드로
arXiv cs.AI · In-Chang Baek, Seoyoung Lee, Sung-Hyun Kim, Geumhwan Hwang, KyungJoong Kim · 2026-07-20 AI

[Submitted on 13 Aug 2025 (v1), last revised 17 Jul 2026 (this version, v2)]

View PDF HTML (experimental)

Abstract:Human-aligned AI is a critical component of co-creativity, as it enables models to accurately interpret human intent and generate controllable outputs that align with design goals in collaborative content creation. This direction is especially relevant in procedural content generation via reinforcement learning (PCGRL), which is intended to serve as a tool for human designers. However, existing systems often fall short of exhibiting human-centered behavior, limiting the practical utility of AI-driven generation tools in real-world design workflows. In this paper, we propose VIPCGRL (Vision-Instruction PCGRL), a novel deep reinforcement learning framework that incorporates three modalities-text, level, and sketches-to extend control modality and enhance human-likeness. We introduce a shared embedding space trained via quadruple contrastive learning across modalities and human-AI styles, and align the policy using an auxiliary reward based on embedding similarity. Experimental results show that VIPCGRL outperforms existing baselines in human-likeness, as validated by both quantitative metrics and human evaluations. The code and dataset are available at this https URL.

Submission history

From: In-Chang Baek [view email]
[v1] Wed, 13 Aug 2025 14:52:14 UTC (12,188 KB)
[v2] Fri, 17 Jul 2026 03:52:55 UTC (9,669 KB)

원문에서 계속 ↗

추출 본문 · 출처: arxiv.org · https://arxiv.org/abs/2508.09860

코멘트

답글 남기기

이메일 주소는 공개되지 않습니다. 필수 필드는 *로 표시됩니다