arXiv:2507.02941cs.CVcs.AI2025-07中稿 · oral presentation …被引 3

构建低分辨率游戏贴图语义数据集,助力叙事驱动的自动内容生成。

GameTileNet: A Semantic Dataset for Low-Resolution Game Art in Procedural Content Generation

  • 收集开源艺术贴图并标注语义、连接关系与物体分类
  • 支持32x32像素低分辨率贴图中的目标检测任务
  • 适合游戏开发与生成式AI研究者使用

GameTileNet 是一个为低分辨率数字游戏艺术设计的语义标注数据集,推动程序化内容生成(PCG)及相关AI研究,作为视觉-语言对齐任务。大型语言模型(LLMs)和图像生成AI已帮助独立开发者生成游戏交互用的精灵图等视觉资源,但生成结果在叙事一致性上仍面临挑战,需人工调整。此外,因训练数据风格分布不均,自动生成内容的视觉多样性受限。GameTileNet通过采集 OpenGameArt.org 上符合创作共用许可的艺术贴图,并提供语义标注,支持叙事驱动的内容生成。该数据集引入了低分辨率贴图(如32x32像素)中的目标检测流程,标注了语义、连通性与物体类别。它是改进PCG方法、实现叙事丰富游戏内容的重要资源,并为低分辨率非写实图像的目标检测提供了基准。简言之,GameTileNet 是一个用于叙事驱动程序化内容生成的低分辨率游戏贴图语义数据集。

原文摘要 · Abstract (English)

GameTileNet is a dataset designed to provide semantic labels for low-resolution digital game art, advancing procedural content generation (PCG) and related AI research as a vision-language alignment task. Large Language Models (LLMs) and image-generative AI models have enabled indie developers to create visual assets, such as sprites, for game interactions. However, generating visuals that align with game narratives remains challenging due to inconsistent AI outputs, requiring manual adjustments by human artists. The diversity of visual representations in automatically generated game content is also limited because of the imbalance in distributions across styles for training data. GameTileNet addresses this by collecting artist-created game tiles from OpenGameArt.org under Creative Commons licenses and providing semantic annotations to support narrative-driven content generation. The dataset introduces a pipeline for object detection in low-resolution tile-based game art (e.g., 32x32 pixels) and annotates semantics, connectivity, and object classifications. GameTileNet is a valuable resource for improving PCG methods, supporting narrative-rich game content, and establishing a baseline for object detection in low-resolution, non-photorealistic images. TL;DR: GameTileNet is a semantic dataset of low-resolution game tiles designed to support narrative-driven procedural content generation through visual-language alignment.

程序化生成游戏艺术语义标注低分辨率

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。