arrow
Return

ScribbleSense: Generative Scribble-Based Texture Editing With Intent Prediction

delete2025-11-21
delete0
PRE
AI
Y
Yudi Zhang
Y
Yeming Geng
张磊 cover
张磊 (Lei Zhang)
DOI:10.1109/TVCG.2025.3635035delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Interactive 3D model texture editing presents enhanced opportunities for creating 3D assets, with freehand drawing style offering the most intuitive experience. However, existing methods primarily support sketch-based interactions for outlining, while the utilization of coarse-grained scribble-based interaction remains limited. Furthermore, current methodologies often encounter challenges due to the abstract nature of scribble instructions, which can result in ambiguous editing intentions and unclear target semantic locations. To address these issues, we propose ScribbleSense, an editing method that combines multimodal large language models (MLLMs) and image generation models to effectively resolve these challenges. We leverage the visual capabilities of MLLMs to predict the editing intent behind the scribbles. Once the semantic intent of the scribble is discerned, we employ globally generated images to extract local texture details, thereby anchoring local semantics and alleviating ambiguities concerning the target semantic locations. Experimental results indicate that our method effectively leverages the strengths of MLLMs, achieving state-of-the-art interactive editing performance for scribble-based texture editing.
Keywords:
Texture editing
large-language models
diffusion models
3D textured meshes

Journal

IEEE Transactions on Visualization and Computer Graphics cover
IEEE Transactions on Visualization and Computer Graphics
IF:
6.5
Papers:
337
Citations:
2.2W

Organization

B
beijing institute of technology
Scholars:
5.5W
Papers: 4.0W
Citations: 63
Cited Papers

Cited Papers

Instructive3D: Editing Large Reconstruction Models with Text Instructions
err2025-02-26
err0
PREAI
errKathare,Kunal; Dhiman,Ankit; Gowda,K Vikas; Aravindan,Siddharth; Monga,Shubham; Vandrotti,Basavaraja Shanthappa; Boregowda,Lokesh R
errShare
errSave
Paint3D: Paint Anything 3D With Lighting-Less Texture Diffusion Models
err2024-06-16
err0
PREAI
errXianfang Zeng; Xin Chen; Zhongqi Qi; Wen Liu; Zibo Zhao; Zhibin Wang; Bin Fu; Yong Liu; Gang Yu
errShare
errSave
TEXGen: a Generative Diffusion Model for Mesh Textures
err2024-11-19
err0
PREAI
errYu, Xin; Yuan, Ze; Guo, Yuan-chen; Liu, Ying-tian; Liu, Jian hui; Li, Yangguang; Cao, Yan-pei; Liang, Ding; Qi, Xiaojuan
errShare
errSave
High-Resolution Image Synthesis with Latent Diffusion Models
err2022-06-01
err0
errOAAI
errRobin Rombach; Andreas Blattmann; Dominik Lorenz; Patrick Esser; Bjorn Ommer
errShare
errSave
Pixel Aligned Language Models
err2024-06-16
err0
PREAI
errXu,Jiarui; Zhou,Xingyi; Yan,Shen; Gu,Xiuye; Arnab,Anurag; Sun,Chen; Wang,Xiaolong; Schmid,Cordelia
errShare
errSave
Recognize Anything: A Strong Image Tagging Model
err2024-06-17
err0
PREAI
errYoucai Zhang; Xinyu Huang; Jinyu Ma; Zhaoyang Li; Zhaochuan Luo; Yanchun Xie; Yuzhuo Qin; Tong Luo; Yaqian Li; Shilong Liu; Yandong Guo; Lei Zhang
errShare
errSave
Objaverse: A Universe of Annotated 3D Objects
err2023-06-01
err0
errOAAI
errMatt Deitke; Dustin Schwenk; Jordi Salvador; Luca Weihs; Oscar Michel; Eli VanderBilt; Ludwig Schmidt; Kiana Ehsanit; Aniruddha Kembhavi; Ali Farhadi
errShare
errSave
errShare
errSave
errShare
errSave
researcher View more