
Identifies a class of prompts that pass text filters yet still produce unsafe images, and defends against it purely at the prompt layer — usable against closed image models where weight editing or encoder tuning is impossible.
articleBeyond Prompt Engineering: A Systematic Analysis of Prompt Lexical Sensitivity and Its Impacts on QualityQipeng Xie, Zi Liang, Jiafei Wu, Yufei Chen, Weizheng Wang, Wenao Ma, Zhong Ming, Haiqin Yang, Kaishun Wu
articleSafety for Whom? Refusing the Right Subset of a Topic, Not the Whole Topichuggingface.co
articleMitigating Stereotypical Biases In Text To Image Generative SystemsRunwayChecking sign-in…
Loading comments…