用谷歌Gemini生成6291张多语言科学插图,支持科研可视化研究。
SciDraw-6K: A Multilingual Scientific Illustration Dataset Generated by Google Gemini

- 用Gemini模型生成跨语言科学插图,覆盖8大领域。
- 数据集含6291张图,每张配11种语言提示,用于精准生成。
- 适合做多语言科学绘图、扩散模型微调和提示工程研究。
我们提出SciDraw-6K,一个由Google Gemini图像生成模型合成的科学插图数据集,共包含6,291张图像,每张均配有11种语言(英语、简体中文、繁体中文、日语、韩语、德语、法语、西班牙语、巴西葡萄牙语、意大利语、俄语)的提示词。图像涵盖生物医学、化学、材料、电子、环境、人工智能系统、物理以及一个“其他”长尾类别,主要由gemini-2.5-flash-image和gemini-3-pro-image-preview模型生成。与主流通用文本到图像数据集不同,SciDraw-6K专为科学插图设计,包括示意图、机制图、目录图和概念海报等类型。本文描述了构建流程,报告了数据统计,并展示了其在sci-draw.com公开绘图服务中的应用。该数据集可用于多语言文本到图像研究、领域自适应扩散模型微调及科学可视化提示工程。数据集地址:https://huggingface.co/datasets/SciDrawAI/SciDraw-6K;代码地址:https://github.com/SciDrawAI/scidraw-6k
原文摘要 · Abstract (English)
We present SciDraw-6K, a curated dataset of 6,291 scientific illustrations synthesized by Google Gemini image-generation models, each paired with prompts in eleven languages (English, Simplified Chinese, Traditional Chinese, Japanese, Korean, German, French, Spanish, Brazilian Portuguese, Italian, and Russian). Images span eight broad scientific categories -- biomedical, chemistry, materials, electronics, environment, AI systems, physics, and a long "other" tail -- and are produced primarily by the gemini-2.5-flash-image and gemini-3-pro-image-preview model families. In contrast to general-purpose text-to-image corpora that dominate the literature, SciDraw-6K is purpose-built for the scientific illustration genre: schematic diagrams, mechanism figures, table-of-contents graphics, and conceptual posters. We describe the construction pipeline, report dataset statistics, and document its use as the substrate of sci-draw.com, a public scientific drawing service. The dataset is released to support multilingual text-to-image research, domain-adapted diffusion fine-tuning, and prompt-engineering studies for scientific visualization. Dataset: https://huggingface.co/datasets/SciDrawAI/SciDraw-6K Code: https://github.com/SciDrawAI/scidraw-6k
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。