返回
Controllable diffusion models for hazardous construction site scene generation
DOI:10.1016/j.asoc.2025.113446.png)
摘要
En 中文
• 两阶段LLM-diffusion框架生成施工危险场景。
• 布局模块对齐布局文本注意力以实现精确图像控制。
• 尺度融合优化U-Net特征以提升生成质量。
• 实验表明在危险场景中具有更优的可控性和质量。
期刊
IF:
6.6
论文数:
1.4W
被引数:
4.8W
机构
暂无机构信息
引用论文
Automated annotation for visual recognition of construction resources using synthetic images使用合成图像对建筑资源进行视觉识别的自动注释
Image generation of hazardous situations in construction sites using text-to-image generative model for training deep neural networks使用用于训练深度神经网络的文本到图像生成模型生成建筑工地危险情况的图像
没有更多内容

