用AI生成超现实主义风格画作,发现DALL-E 2结合ChatGPT提示效果最佳。
Surrealistic-like Image Generation with Vision-Language Models
- 用ChatGPT生成提示词,配合DALL-E 2生成超现实图像。
- 实验表明DALL-E 2在各类模型中生成质量最高。
- 适合对艺术生成和提示工程感兴趣的开发者或艺术家。
近期生成式AI的发展使文本、图像、代码等内容的生成变得便捷。本文探索使用视觉语言生成模型(包括DALL-E、Deep Dream Generator和DreamStudio)生成超现实主义风格画作的可能性。研究从不同生成设置和模型出发,旨在识别最适合生成此类图像的模型与参数配置,并分析使用编辑过的原始图像对生成结果的影响。通过实验评估选定模型的表现,获得其在生成超现实风格图像方面的关键洞察。结果表明,在使用ChatGPT生成的提示词下,DALL-E 2表现最优。
原文摘要 · Abstract (English)
Recent advances in generative AI make it convenient to create different types of content, including text, images, and code. In this paper, we explore the generation of images in the style of paintings in the surrealism movement using vision-language generative models, including DALL-E, Deep Dream Generator, and DreamStudio. Our investigation starts with the generation of images under various image generation settings and different models. The primary objective is to identify the most suitable model and settings for producing such images. Additionally, we aim to understand the impact of using edited base images on the generated resulting images. Through these experiments, we evaluate the performance of selected models and gain valuable insights into their capabilities in generating such images. Our analysis shows that Dall-E 2 performs the best when using the generated prompt by ChatGPT.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。