[2602.04898] Semantic-level Backdoor Attack against Text-to-Image

[2602.04898] Semantic-level Backdoor Attack against Text-to-Image Diffusion Models

arXiv - AI March 04, 2026 3 min read

About this article

Abstract page for arXiv paper 2602.04898: Semantic-level Backdoor Attack against Text-to-Image Diffusion Models

Computer Science > Cryptography and Security arXiv:2602.04898 (cs) [Submitted on 3 Feb 2026 (v1), last revised 3 Mar 2026 (this version, v2)] Title:Semantic-level Backdoor Attack against Text-to-Image Diffusion Models Authors:Tianxin Chen, Wenbo Jiang, Hongqiao Chen, Zhirun Zheng, Cheng Huang View a PDF of the paper titled Semantic-level Backdoor Attack against Text-to-Image Diffusion Models, by Tianxin Chen and 4 other authors View PDF HTML (experimental) Abstract:Text-to-image (T2I) diffusion models are widely adopted for their strong generative capabilities, yet remain vulnerable to backdoor attacks. Existing attacks typically rely on fixed textual triggers and single-entity backdoor targets, making them highly susceptible to enumeration-based input defenses and attention-consistency detection. In this work, we propose Semantic-level Backdoor Attack (SemBD), which implants backdoors at the representation level by defining triggers as continuous semantic regions rather than discrete textual patterns. Concretely, SemBD injects semantic backdoors by distillation-based editing of the key and value projection matrices in cross-attention layers, enabling diverse prompts with identical semantic compositions to reliably activate the backdoor attack. To further enhance stealthiness, SemBD incorporates a semantic regularization to prevent unintended activation under incomplete semantics, as well as multi-entity backdoor targets that avoid highly consistent cross-attention pattern...

Originally published on March 04, 2026. Curated by AI News.

Machine Learning

Making an AI native sovereign computational stack

I’ve been working on a personal project that ended up becoming a kind of full computing stack: identity / trust protocol decentralized ch...

Reddit - Artificial Intelligence · 1 min · 32 minutes ago

Llms

An attack class that passes every current LLM filter - no payload, no injection signature, no log trace

https://shapingrooms.com/research I published a paper today on something I've been calling postural manipulation. The short version: ordi...

Reddit - Artificial Intelligence · 1 min · 32 minutes ago

Machine Learning

What tools are sr MLEs using? (clawdbot, openspec, wispr) [D]

I'm already blasting cursor, but I want to level up my output. I heard that these kind of AI tools and workflows are being asked in SF. W...

Reddit - Machine Learning · 1 min · about 1 hour ago

Machine Learning

[R] looking for academic collaborators

hey there, i am currently working with a research group at auckland university. we are currently working on neurodegenerative diseases - ...

Reddit - Machine Learning · 1 min · about 1 hour ago

[2602.04898] Semantic-level Backdoor Attack against Text-to-Image Diffusion Models

About this article

Related Articles

Making an AI native sovereign computational stack

An attack class that passes every current LLM filter - no payload, no injection signature, no log trace

What tools are sr MLEs using? (clawdbot, openspec, wispr) [D]

[R] looking for academic collaborators

No comments

Stay updated with AI News