REVIEW 2 major objections 6 minor 2 cited by
Inkspire: Supporting Design Exploration with Generative AI through Analogical Sketching
T0 review · 2 major / 6 minor · reviewed 2026-08-09 · deepseek-v4-flash
Pith's one-line read Inkspire claims a sketch-to-design-to-sketch loop with analogical inspiration helps designers explore more and avoid AI fixation.
desk verdict Solid systems paper with a confounded comparison: the workflow is new and the effects are large, but the specific claim about analogical sketching isn't uniquely supported by the data. read the letter →
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
The reading
What carries the argument
Three coupled mechanisms carry the argument. First, analogical inspiration: an LLM prompted with chain-of-thought reasoning converts an abstract concept into ten visually concrete objects drawn from nature, architecture, and fashion, giving designers concept-level anchors without prompt engineering. Second, per-stroke generation with a dynamic guidance scale G(n)=7-4·0.$5^{{n/3}}$, which starts near 3 when the canvas is nearly empty and asymptotes at 7 as strokes accumulate, so ControlNet can interpret incomplete sketches and re-render after every pen stroke while keeping the seed fixed for continuity. Third, the Design2Sketch pipeline converts the high-fidelity generation into a scaffold via Scaffolding = Boundary(Seg(D)) ∩ SoftEdge(D): semantic segmentation boundaries are intersected with HED soft edges to keep only key structural lines, which are then shown as a tracing-paper-style underlay. The scaffold is the hinge of the loop—it is the mechanism that turns a 'too complete' output back into something malleable.
What would settle it
A controlled experiment that augments the ControlNet baseline with automatic per-stroke regeneration and sketch scaffolding but not the analogical inspiration panel would locate the source of the exploration gain. If that augmented baseline shows the same exploration and inspiration ratings as Inkspire, the central claim that analogical sketching drives the effect is falsified. Alternatively, an objective divergence measure—counting the number of structurally distinct concepts in the final designs produced by each condition—would test whether self-reported exploration translates into more varied outcomes.
Extended reading notes
Core claim
The central claim is that a complete feedback loop—sketch guides AI generation, AI output is abstracted back into a sketch scaffold, and the scaffold guides the next stroke—supports a more iterative, exploratory, and co-creative design workflow than the current practice of sketching a full image and handing it to a ControlNet model with a text prompt. The paper reports that the twelve participants rated Inkspire significantly higher than the baseline on inspiration (t(11)=3.44, p<0.01) and exploration (t(11)=3.94, p<0.01), as well as on controllability, communication, partnership, and attribution (all p<0.01). Interaction logs show Inkspire users sketching in short bursts with frequent generations, while baseline users drew long sequences and edited prompts incrementally; prompt semantic similarity was lower under Inkspire (BERTScore 0.51 vs. 0.76). The authors attribute these differences to three mechanisms: analogical inspiration that turns abstract briefs into concrete visual anchors, per-stroke regeneration with a dynamically increasing guidance scale that tolerates incomplete sketches, and sketch scaffolding that lets designers build on a generation without being fixated on its photorealistic finish.
Load-bearing premise
The baseline ControlNet condition differs from Inkspire on several dimensions at once—analogical inspirations, sketch scaffolding, per-stroke regeneration, and the dynamic guidance scale—so the reported benefits may stem from any subset of these differences rather than from the complete loop, and the paper does not include an ablation to isolate them.
Editorial extensions
If this is right
- Designers can start ideation from a single abstract word and a single stroke rather than a fully specified prompt or sketch.
- Per-stroke regeneration creates a turn-taking rhythm that makes the AI feel like a collaborator rather than a one-shot renderer.
- The scaffolding underlay could be applied to other generative domains where high-fidelity outputs cause fixation.
- The dynamic guidance scale provides a general recipe for making ControlNet-style models tolerate partial input.
Reading between the lines
- If the analogical menu is the driver, a prompt-based tool with analogies but no scaffolding might reproduce part of the exploration gain, suggesting a possible ablation.
- The measured reduction in prompt editing and the lower BERTScore similarity suggest that the interface substitutes conceptual pivots for lexical tweaks; whether this produces objectively more novel final designs is not fully settled by self-reports.
- The single-thread limitation noted in the paper implies the loop may benefit from parallel analogy branches, which the authors themselves flag as future work.
- Beyond product design, the sketch-to-design-to-sketch loop could be adapted to architecture or fashion sketching if the analogical source domains are re-targeted.
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. This paper presents Inkspire, a sketch-driven text-to-image tool that combines an analogical inspiration panel (LLM-generated concrete analogies for abstract concepts), a sketching canvas with per-stroke auto-regeneration using a dynamic guidance scale, and a Design2Sketch pipeline that converts AI-generated designs into sketch-style underlays. The authors report a within-subjects study with twelve participants comparing Inkspire to a ControlNet-based baseline, finding significantly higher self-reported exploration, inspiration, controllability, communication, partnership, and attribution, and concluding that Inkspire promotes a more iterative, exploratory, and co-creative workflow that helps designers overcome fixation.
Significance. If the results hold, Inkspire is a meaningful step toward reducing design fixation in T2I workflows, offering a concrete system with a novel Design2Sketch pipeline and a validated evaluation using standard questionnaires. The paper reports large effect sizes, transparently presents non-significant results, and grounds its design goals in a professional design team exchange. However, the significance rests on the strength of the comparative evidence, which is currently confounded by multiple simultaneous differences between the conditions; therefore the precise contribution of analogical sketching is not yet established.
major comments (2)
- [§5.3, §6.1, §7.1] The central evaluation is a system-level comparison in which Inkspire differs from the baseline on at least four dimensions at once: the analogical inspiration menu, the sketch-scaffold underlay, automatic per-stroke regeneration with the dynamic guidance scale (Eq. 1), and the baseline's requirement to complete a full sketch before generating. The significant gains in exploration (t(11)=3.94) and inspiration (t(11)=3.44) could therefore be driven by the per-stroke feedback loop alone rather than by analogical sketching. Because the abstract and conclusion attribute the outcome to 'analogical sketching' and the sketch-to-design-to-sketch loop, this confound is load-bearing. Section 7.1 correctly lists ablation studies as future work, but the claims as written go beyond what the current comparison can establish. The authors should either narrow the central claim to the integrated system or add an analytic or experimental disaggregation (e.g., a condition without the analogy panel, or a baseline with per-stroke generation) to support the current wording.
- [§6.2.3] The logged sketching behavior reveals a strong asymmetry: baseline users averaged 59.8 strokes before generating, while Inkspire users averaged 17.3 strokes and interleaved one or a few strokes between generations. The baseline's full-sketch-before-generation requirement changes the cost of exploration, so the observed self-reported benefits may reflect this interaction economics rather than the value of analogies or scaffolds. To make the baseline a fair point of comparison, the authors need to justify that this requirement is representative of current practice, and ideally include a condition that gives ControlNet a comparable per-stroke interaction (as suggested in §7.1). Without such a control, the conclusion that analogical sketching is the active ingredient is not uniquely supported.
minor comments (6)
- [§4.2, Eq. (1)] The dynamic guidance scale is a hand-selected formula; a brief sensitivity discussion or reference to tuning experiments would help readers understand why the constants were chosen.
- [§6.2.2] The BERTScore semantic-similarity comparison is reported as 'much lower' without a test statistic or p-value; a paired t-test or equivalent should be reported.
- [§6.2.3] The claim that participants drew fewer total strokes with Inkspire is not accompanied by a significance test; if the difference is not tested, it should be described as an observation only.
- [Table 1] There appear to be inconsistencies between the listed analogies and the 'Total' count for several participants (e.g., P1 lists seven items but reports six; P4 lists five items but reports four). Please check and correct.
- [§4.3, §2.2] There is a typo in §4.3 ('acheve' should be 'achieve') and a doubled 'and' in §2.2 ('LLMs and and analogical reasoning').
- [§3] The formative exchange session is described as a day-long session with seven designers, but no details on the protocol or analysis are provided; adding a brief description would strengthen the derivation of the design goals.
Circularity Check
No significant circularity; the user-study evaluation is an empirical comparison rather than a derivation from fitted inputs.
full rationale
The paper does not claim to derive quantitative predictions from fitted parameters. Equation (1), the dynamic guidance scale G(n)=7-4*0.5^(n/3), is a hand-selected design rule for enabling per-stroke generation, and Equation (2), the Design2Sketch scaffolding, is a definitional composition of existing semantic segmentation and soft-edge extraction methods. Neither equation is calibrated to the study outcomes nor used to predict the reported exploration, inspiration, or collaboration ratings. The central evidence is a within-subjects empirical comparison against a ControlNet baseline, with the headline results reported as paired t-tests on self-report scales and logged interaction behavior. Those results are not reductions of outputs to inputs by construction. Self-citations such as BioSpark [39] and Jigsaw [51] appear in related-work, motivation, and future-work contexts, and are not invoked as load-bearing uniqueness theorems or used to forbid alternative explanations. The paper explicitly acknowledges in Section 7.1 that Inkspire contains multiple features that could affect behavior and that ablation studies are future work; this is a confound/internal-validity limitation, not a circular derivation. Therefore the paper is self-contained with respect to circularity, and any concerns about the strength of the baseline belong under correctness risk rather than circularity.
Assumptions & free parameters
free parameters (2)
- Guidance scale constants =
G(n)=7-4*0.5^(n/3)
- Source domains for analogies =
nature, architecture, fashion
assumptions (5)
- domain assumption The formative session with seven automotive designers yields design goals that generalize to the study participants and broader product designers.
- domain assumption Lower-fidelity sketch scaffolds reduce design fixation, based on prior work.
- domain assumption GPT-4 produces useful, unbiased analogical inspirations suitable for product design.
- standard math Parametric paired t-tests are valid for 7-point Likert questionnaire data with n=12.
- domain assumption The baseline ControlNet interface is a fair comparison point.
Cite this review
Pith. "Pith review of Inkspire: Supporting Design Exploration with Generative AI through Analogical Sketching." pith.science (2026). https://pith.science/paper/XUURMDJI
@misc{pith2026250118588,
author = {Pith},
title = {Pith review of: Inkspire: Supporting Design Exploration with Generative AI through Analogical Sketching},
year = {2026},
howpublished = {\url{https://pith.science/paper/XUURMDJI}},
note = {Machine review of arXiv:2501.18588}
}
read the original abstract
With recent advancements in the capabilities of Text-to-Image (T2I) AI models, product designers have begun experimenting with them in their work. However, T2I models struggle to interpret abstract language and the current user experience of T2I tools can induce design fixation rather than a more iterative, exploratory process. To address these challenges, we developed Inkspire, a sketch-driven tool that supports designers in prototyping product design concepts with analogical inspirations and a complete sketch-to-design-to-sketch feedback loop. To inform the design of Inkspire, we conducted an exchange session with designers and distilled design goals for improving T2I interactions. In a within-subjects study comparing Inkspire to ControlNet, we found that Inkspire supported designers with more inspiration and exploration of design ideas, and improved aspects of the co-creative process by allowing designers to effectively grasp the current state of the AI to guide it towards novel design intentions.
Figures
Figures from the paper (11 more)
Forward citations
Cited by 2 Pith papers
-
Canvas3D: Empowering Precise Spatial Control for Image Generation with Constraints from a 3D Virtual Canvas
Canvas3D lets users arrange objects in a 3D canvas generated from a text prompt, then feeds depth, skeleton, and lighting constraints to diffusion models to produce images that match the layout.
-
GenTune: Toward Traceable Prompts to Improve Controllability of Image Refinement in Environment Design
GenTune improves AI image refinement by tracing image regions back to prompt labels and allowing element-level, semantic-guided edits.
Reference graph
Works this paper leans on
-
[1]
2022. Upwork. Retrieved August 15, 2022 from https://www.upwork.com/
work page 2022
-
[2]
AI is plundering the imagination and replacing it with a slot machine
2024. AI is plundering the imagination and replacing it with a slot machine . Retrieved March 24, 2024 from https://thebulletin.org/2022/10/ai-is-plundering- the-imagination-and-replacing-it-with-a-slot-machine/
work page 2024
-
[3]
How to Learn to Draw by Tracing
2024. How to Learn to Draw by Tracing . Retrieved February 16, 2024 from https://monikazagrobelna.com/2020/08/16/how-to-learn-to-draw-by-tracing/
work page 2024
- [4]
-
[5]
2024. Vizcom. Retrieved March 24, 2024 from https://app.vizcom.ai/
work page 2024
-
[6]
Barrett R Anderson, Jash Hemant Shah, and Max Kreminski. 2024. Homog- enization Effects of Large Language Models on Human Creative Ideation. In Proceedings of the 16th Conference on Creativity & Cognition (Chicago, IL, USA) (C&C ’24). Association for Computing Machinery, New York, NY, USA, 413–425. https://doi.org/10.1145/3635636.3656204
-
[7]
Hal R Arkes and Catherine Blumer. 1985. The psychology of sunk cost. Organi- zational behavior and human decision processes 35, 1 (1985), 124–140
work page 1985
-
[8]
Luca Benedetti, Holger Winnemöller, Massimiliano Corsini, and Roberto Scopigno. 2014. Painting with Bob: assisted creativity for novices. In Proceedings of the 27th annual ACM symposium on User interface software and technology . 419–428
work page 2014
Show all 78 references
-
[9]
James Betker, Gabriel Goh, Li Jing, Tim Brooks, Jianfeng Wang, Linjie Li, Long Ouyang, Juntang Zhuang, Joyce Lee, Yufei Guo, et al . 2023. Improving im- age generation with better captions. Computer Science. https://cdn. openai. com/papers/dall-e-3. pdf 2, 3 (2023), 8
2023
-
[10]
Antoine Bordas, Pascal Le Masson, and Benoit Weil. 2024. Switching perspectives on generative artificial intelligence: a design view for humans-generative AI co-creativity. In R&D Management Conference 2024 . 14
2024
-
[11]
Stephen Brade, Bryan Wang, Mauricio Sousa, Sageev Oore, and Tovi Gross- man. 2023. Promptify: Text-to-image generation through interactive prompt exploration with large language models. In Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology . 1–14
2023
-
[12]
Tim Brooks, Aleksander Holynski, and Alexei A Efros. 2023. Instructpix2pix: Learning to follow image editing instructions. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 18392–18402
2023
-
[13]
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. 2020. Language models are few-shot learners. Advances in neural information processing systems 33 (2020), 1877–1901
2020
-
[14]
Raluca Budiu. 2014. Memory recognition and recall in user interfaces. Nielsen Norman Group 1 (2014)
2014
-
[15]
Bill Buxton. 2010. Sketching user experiences: getting the design right and the right design. Morgan kaufmann
2010
-
[16]
Alice Cai, Steven R Rick, Jennifer L Heyman, Yanxia Zhang, Alexandre Filipowicz, Matthew Hong, Matt Klenk, and Thomas Malone. 2023. DesignAID: Using Generative AI and Semantic Diversity for Design Inspiration. InProceedings of The ACM Collective Intelligence Conference (Delft,...
2023
-
[17]
John Canny. 1986. A computational approach to edge detection.IEEE Transactions on pattern analysis and machine intelligence 6 (1986), 679–698
1986
-
[18]
Carlos Cardoso, Petra Badke-Schaub, and Ana Luz. 2009. Design fixation on non- verbal stimuli: The influence of simple vs. rich pictorial information on design problem-solving. In International Design Engineering Technical Conferences and Computers and Information in Engineeri...
2009
-
[19]
How To Draw Cars. 2017. Pro Designer Teaches You How To Create Varia- tions on Your Car Design Themes. https://youtu.be/GxFloTllsKE?si=yWYl9w_ etZVZPbnK&t=80. Accessed: September 12, 2024
2017
-
[20]
Minsuk Chang, Stefania Druga, Alexander J Fiannaca, Pedro Vergani, Chinmay Kulkarni, Carrie J Cai, and Michael Terry. 2023. The prompt artists. InProceedings of the 15th Conference on Creativity and Cognition . 75–87
2023
-
[21]
Peiyao Cheng, Ruth Mugge, and Jan PL Schoormans. 2014. A new strategy to reduce design fixation: Presenting partial photographs to designers. Design Studies 35, 4 (2014), 374–391
2014
-
[22]
Erin Cherry and Celine Latulipe. 2014. Quantifying the creativity support of digital tools through the creativity support index.ACM Transactions on Computer- Human Interaction (TOCHI) 21, 4 (2014), 1–25
2014
-
[23]
Nicholas Davis, Safat Siddiqui, Pegah Karimi, Mary Lou Maher, and Kazjon Grace
-
[25]
de Rooij and M
A. de Rooij and M. Mose Biskjaer. 2024. Expecting the unexpected: A review of surprise in design processes. In DRS2024: Boston, 23–28 June , C. Gray, E. Cil- iotta Chehade, P. Hekkert, L. Forlano, P. Ciuccarelli, and P. Lloyd (Eds.). Boston, USA. https://doi.org/10.21606/drs.2024.333
2024 doi
-
[26]
Jon-Michael Deldin and Megan Schuknecht. 2013. The AskNature database: enabling solutions in biomimetic design. In Biologically inspired design: Compu- tational methods and tools . Springer, 17–27
2013
-
[27]
Steven P Dow, Alana Glassco, Jonathan Kass, Melissa Schwarz, Daniel L Schwartz, and Scott R Klemmer. 2010. Parallel prototyping leads to better design results, more divergence, and increased self-efficacy. ACM Transactions on Computer- Human Interaction (TOCHI) 17, 4 (2010), 1–24
2010
-
[28]
Zezhong Fan, Xiaohan Li, Kaushiki Nag, Chenhao Fang, Topojoy Biswas, Jianpeng Xu, and Kannan Achan. 2024. Prompt Optimizer of Text-to-Image Diffusion Models for Abstract Concept Understanding. In Companion Proceedings of the ACM Web Conference 2024 (Singapore, Singapore) (WWW ...
2024
-
[29]
Dedre Gentner. 1983. Structure-mapping: A theoretical framework for analogy. Cognitive science 7, 2 (1983), 155–170
1983
-
[30]
Ashok K Goel. 1997. Design, analogy, and creativity. IEEE expert 12, 3 (1997), 62–70
1997
-
[31]
Ashok K Goel, Swaroop Vattam, Bryan Wiltgen, and Michael Helms. 2012. Cog- nitive, collaborative, conceptual and creative—Four characteristics of the next generation of knowledge-based CAD systems: A study in biologically inspired design. Computer-Aided Design 44, 10 (2012), 879–900
2012
-
[32]
Yihan Hou, Manling Yang, Hao Cui, Lei Wang, Jie Xu, and Wei Zeng. 2024. C2Ideas: Supporting Creative Interior Color Design Ideation with Large Language Model. arXiv preprint arXiv:2401.12586 (2024)
2024 arXiv
-
[33]
Zhengyu Huang, Yichen Peng, Tomohiro Hibino, Chunqi Zhao, Haoran Xie, Tsukasa Fukusato, and Kazunori Miyata. 2022. dualface: Two-stage drawing guidance for freehand portrait sketching. Computational Visual Media 8 (2022), 63–77
2022
-
[34]
Emmanuel Iarussi, Adrien Bousseau, and Theophanis Tsandilas. 2013. The draw- ing assistant: Automated drawing guidance and feedback from photographs. In ACM Symposium on User Interface Software and Technology (UIST) . ACM
2013
-
[35]
David G Jansson and Steven M Smith. 1991. Design fixation. Design studies 12, 1 (1991), 3–11
1991
-
[36]
Shuo Jiang, Jie Hu, Kristin L Wood, and Jianxi Luo. 2022. Data-driven design-by- analogy: state-of-the-art and future directions. Journal of Mechanical Design 144, 2 (2022), 020801
2022
-
[37]
Shuo Jiang, Jianxi Luo, Guillermo Ruiz-Pava, Jie Hu, and Christopher L Magee
-
[38]
Caneel K Joyce. 2009. The blank page: Effects of constraint on creativity. University of California, Berkeley
2009
-
[39]
Hyeonsu B Kang, David Chuan-En Lin, Nikolas Martelaro, Aniket Kittur, Yan- Ying Chen, and Matthew K Hong. 2023. BioSpark: An End-to-End Genera- tive System for Biological-Analogical Inspirations and Ideation. arXiv preprint arXiv:2312.11388 (2023)
2023 arXiv
-
[40]
Mohammed Khaliq, Diego Frassinelli, and Sabine Schulte im Walde. 2024. Compar- ison of Image Generation Models for Abstract and Concrete Event Descriptions. In Proceedings of the 4th Workshop on Figurative Language Processing (FigLang 2024). 15–21
2024
-
[41]
Alexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao, Chloe Rolland, Laura Gustafson, Tete Xiao, Spencer Whitehead, Alexander C Berg, Wan-Yen Lo, et al
-
[42]
Ken Kocienda. 2018. Creative selection: Inside Apple’s design process during the golden age of Steve Jobs . Pan Macmillan
2018
-
[43]
E Kwon, A Pehlken, K-D Thoben, A Bazylak, and LH Shu. 2019. Visual similarity to aid alternative-use concept generation for retired wind-turbine blades.Journal of Mechanical Design 141, 3 (2019), 031106
2019
-
[44]
Tomas Lawton, Francisco J Ibarrola, Dan Ventura, and Kazjon Grace. 2023. Draw- ing with reframer: Emergence and control in co-creative ai. In Proceedings of the 28th International Conference on Intelligent User Interfaces . 264–277
2023
-
[45]
Seung Won Lee, Tae Hee Jo, Semin Jin, Jiin Choi, Kyungwon Yun, Sergio Bromberg, Seonghoon Ban, and Kyung Hoon Hyun. 2024. The Impact of Sketch- guided vs. Prompt-guided 3D Generative AIs on the Design Exploration Process. In Proceedings of the CHI Conference on Human Factors i...
2024
-
[46]
Yong Jae Lee, C Lawrence Zitnick, and Michael F Cohen. 2011. Shadowdraw: real-time user guidance for freehand drawing. ACM Transactions on Graphics (ToG) 30, 4 (2011), 1–10
2011
-
[47]
Patrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich Küttler, Mike Lewis, Wen-tau Yih, Tim Rocktäschel, et al. 2020. Retrieval-augmented generation for knowledge-intensive nlp tasks. Advances in Neural Information Processing S...
2020
-
[48]
Chengze Li, Xueting Liu, and Tien-Tsin Wong. 2017. Deep extraction of manga structural lines. ACM Transactions on Graphics (TOG) 36, 4 (2017), 1–12
2017
-
[49]
Jiayi Liao, Xu Chen, Qiang Fu, Lun Du, Xiangnan He, Xiang Wang, Shi Han, and Dongmei Zhang. 2024. Text-to-image generation for abstract concepts. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 38. 3360–3368
2024
-
[50]
Alex Limpaecher, Nicolas Feltman, Adrien Treuille, and Michael Cohen. 2013. Real-time drawing assistance through crowdsourcing. ACM Transactions on Graphics (TOG) 32, 4 (2013), 1–8
2013
-
[51]
David Chuan-En Lin and Nikolas Martelaro. 2024. Jigsaw: Supporting Designers to Prototype Multimodal Applications by Assembling AI Foundation Models. In Proceedings of the 2024 CHI Conference on Human Factors in Computing Systems
2024
-
[52]
Zhiyu Lin, Upol Ehsan, Rohan Agarwal, Samihan Dani, Vidushi Vashishth, and Mark Riedl. 2023. Beyond Prompts: Exploring the Design Space of Mixed- Initiative Co-Creativity Systems. arXiv preprint arXiv:2305.07465 (2023)
2023 arXiv
-
[53]
J. S. Linsey, A. B. Markman, and K. L. Wood. 2012. Design by Analogy: A Study of the WordTree Method for Problem Re-Representation. Jour- nal of Mechanical Design 134, 4 (04 2012), 041009. https://doi.org/10.1115/1. 4006145 arXiv:https://asmedigitalcollection.asme.org/mechanic...
2012 doi
-
[54]
Vivian Liu, Tao Long, Nathan Raw, and Lydia Chilton. 2023. Generative disco: Text-to-video generation for music visualization. arXiv preprint arXiv:2304.08551 (2023)
2023 arXiv
-
[55]
Xuebin Qin, Zichen Zhang, Chenyang Huang, Masood Dehghan, Osmar R Zaiane, and Martin Jagersand. 2020. U2-Net: Going deeper with nested U-structure for salient object detection. Pattern recognition 106 (2020), 107404
2020
-
[56]
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer. 2022. High-resolution image synthesis with latent diffusion models. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 10684–10695. 15
2022
-
[57]
Chitwan Saharia, William Chan, Huiwen Chang, Chris Lee, Jonathan Ho, Tim Salimans, David Fleet, and Mohammad Norouzi. 2022. Palette: Image-to-image diffusion models. In ACM SIGGRAPH 2022 conference proceedings . 1–10
2022
-
[58]
Patsorn Sangkloy, Nathan Burnell, Cusuh Ham, and James Hays. 2016. The sketchy database: learning to retrieve badly drawn bunnies. ACM Transactions on Graphics (TOG) 35, 4 (2016), 1–12
2016
-
[59]
Vishnu Sarukkai, Lu Yuan, Mia Tang, Maneesh Agrawala, and Kayvon Fatahalian
-
[60]
L Siddharth and Amaresh Chakrabarti. 2018. Evaluating the impact of Idea-Inspire 4.0 on analogical transfer of concepts. Ai Edam 32, 4 (2018), 431–448
2018
-
[61]
Kihoon Son, DaEun Choi, Tae Soo Kim, Young-Ho Kim, and Juho Kim. 2024. Gen- query: Supporting expressive visual search with generative models. InProceedings of the CHI Conference on Human Factors in Computing Systems . 1–19
2024
-
[62]
Julian FV Vincent and Darrell L Mann. 2002. Systematic technology transfer from biology to engineering. Philosophical Transactions of the Royal Society of London. Series A: Mathematical, Physical and Engineering Sciences 360, 1791 (2002), 159–173
2002
-
[63]
Kelly, Saumya Pareek, Qiushi Zhou, and Eduardo Velloso
Samangi Wadinambiarachchi, Ryan M. Kelly, Saumya Pareek, Qiushi Zhou, and Eduardo Velloso. 2024. The Effects of Generative AI on Design Fixation and Divergent Thinking. In Proceedings of the CHI Conference on Human Factors in Computing Systems (Honolulu, HI, USA) (CHI ’24). As...
2024
-
[64]
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al. 2022. Chain-of-thought prompting elicits reasoning in large language models. Advances in Neural Information Processing Systems 35 (2022), 24824–24837
2022
-
[65]
Blake Williford, Samantha Ray, Jung In Koh, Josh Cherian, Paul Taele, and Tracy Hammond. 2023. Exploring Creativity Support for Concept Art Ideation. In Ex- tended Abstracts of the 2023 CHI Conference on Human Factors in Computing Sys- tems (Hamburg, Germany) (CHI EA ’23). Ass...
2023
-
[66]
Shengqiong Wu, Hao Fei, Hanwang Zhang, and Tat-Seng Chua. 2024. Imagine that! abstract-to-intricate text-to-image synthesis with scene graph hallucination diffusion. In Proceedings of the 37th International Conference on Neural Information Processing Systems (New Orleans, LA, ...
2024
-
[67]
Jun Xie, Aaron Hertzmann, Wilmot Li, and Holger Winnemöller. 2014. PortraitS- ketch: Face sketching assistance for novices. In Proceedings of the 27th annual ACM symposium on User interface software and technology . 407–417
2014
-
[68]
Saining Xie and Zhuowen Tu. 2015. Holistically-nested edge detection. In Pro- ceedings of the IEEE international conference on computer vision . 1395–1403
2015
-
[69]
Yutong Xie, Zhaoying Pan, Jinge Ma, Luo Jie, and Qiaozhu Mei. 2023. A prompt log analysis of text-to-image generation systems. In Proceedings of the ACM Web Conference 2023. 3892–3902
2023
-
[70]
JD Zamfirescu-Pereira, Richmond Y Wong, Bjoern Hartmann, and Qian Yang
-
[71]
Chengzhi Zhang, Weijie Wang, Paul Pangaro, Nikolas Martelaro, and Daragh Byrne. 2023. Generative Image AI Using Design Sketches as input: Opportunities and Challenges. In Proceedings of the 15th Conference on Creativity and Cognition (Virtual Event, USA) (C&C ’23). Association...
2023
-
[72]
Lvmin Zhang, Anyi Rao, and Maneesh Agrawala. 2023. Adding conditional con- trol to text-to-image diffusion models. InProceedings of the IEEE/CVF International Conference on Computer Vision . 3836–3847
2023
-
[73]
Tianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q Weinberger, and Yoav Artzi. 2019. Bertscore: Evaluating text generation with bert. arXiv preprint arXiv:1904.09675 (2019)
2019 arXiv
-
[74]
Zijian Zhang and Yan Jin. 2020. An unsupervised deep learning model to discover visual similarity between sketches for visual analogy support. In International design engineering technical conferences and computers and information in engi- neering conference, Vol. 83976. Ameri...
2020
-
[75]
In Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems
Why Johnny can’t prompt: how non-AI experts try (and fail) to design LLM prompts. In Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems. 1–21
2023
-
[2019]
Creative Sketching Partner: A Co-Creative Sketching Tool to Inspire Design Creativity.. In ICCC. 358–359
-
[2021]
Journal of Mechanical Design 143, 6 (2021), 061405
Deriving design feature vectors for patent images using convolutional neural networks. Journal of Mechanical Design 143, 6 (2021), 061405
2021
-
[2023]
arXiv preprint arXiv:2304.02643 (2023)
Segment anything. arXiv preprint arXiv:2304.02643 (2023)
2023 arXiv
-
[2024]
In Proceedings of the 37th Annual ACM Symposium on User Interface Software and Technology
Block and Detail: Scaffolding Sketch-to-Image Generation. In Proceedings of the 37th Annual ACM Symposium on User Interface Software and Technology . 1–13
Reviewed August 9, 2026 · model on record in the stance chip above.
Discussion (0). Continue with ORCID to comment.