diffusiondb poloclub

DiffusionDB is the first large-scale text-to-image prompt dataset. It contains 2 million images generated by Stable Diffusion using prompts and hyperparameters specified by real users. The unprecedented scale and diversity of this human-actuated dataset provide exciting research opportunities in understanding the interplay between prompts and generative models, detecting deepfakes, and designing human-AI interaction tools to help users more easily use these models.

类型
dataset
许可
cc0-1.0
语言
en
下载量
13,400
点赞
651
访问
public
文件
0

标签

  • 文生图
  • 图像数据集
  • 提示词工程
  • 文本生成图像
  • 稳定扩散
  • 扩散模型

摘要

DiffusionDB 是首个大规模文生图提示词数据集,包含由真实用户使用 Stable Diffusion 生成的 1400 万张图像及其对应提示词和超参数(种子、CFG、步数、采样器等)。该数据集规模庞大且多样性高,可用于研究提示词与生成模型之间的交互关系、深度伪造检测,以及设计与生成式AI交互的人机协作工具。数据集提供 2M 和 Large 两个子集,并以图像文件夹与 Parquet 元数据表的形式分发,便于研究者和开发者高效查询和使用。

README

--- layout: default title: Home nav_order: 1 has_children: false annotations_creators: - no-annotation language: - en language_creators: - found license: - cc0-1.0 multilinguality: - multilingual pretty_name: DiffusionDB size_categories: - n>1T source_datasets: - original tags: - stable diffusion -…

查看完整页面 · 查看原文