1,455 results for diffusion

arxiv.org/abs/2502.10465v1

Image Watermarking of Generative Diffusion Models

Embedding watermarks into the output of generative models is essential for establishing copyright and verifiable ownership over the generated content. Emerging diffusion model watermarking methods either embed watermarks in the frequency domain or of...

arxiv.org/abs/2407.17911v1

ReCorD: Reasoning and Correcting Diffusion for HOI Generation

Diffusion models revolutionize image generation by leveraging natural language to guide the creation of multimedia content. Despite significant advancements in such generative models, challenges persist in depicting detailed human-object interactions...

arxiv.org/abs/2602.17664v1

Sink-Aware Pruning for Diffusion Language Models

Diffusion Language Models (DLMs) incur high inference cost due to iterative denoising, motivating efficient pruning. Existing pruning heuristics largely inherited from autoregressive (AR) LLMs, typically preserve attention sink tokens because AR sink...

arxiv.org/abs/2411.18665v3

SpotLight: Shadow-Guided Object Relighting via Diffusion

Recent work has shown that diffusion models can serve as powerful neural rendering engines that can be leveraged for inserting virtual objects into images. However, unlike typical physics-based renderers, these neural rendering engines are limited by...

arxiv.org/abs/2410.14398v3

Dynamic Negative Guidance of Diffusion Models

Negative Prompting (NP) is widely utilized in diffusion models, particularly in text-to-image applications, to prevent the generation of undesired features. In this paper, we show that conventional NP is limited by the assumption of a constant guidan...

arxiv.org/abs/2311.17175v3

Kicking it Off(-shell) with Direct Diffusion

Off-shell effects in large LHC backgrounds are crucial for precision predictions and, at the same time, challenging to simulate. We present a novel method to transform high-dimensional distributions based on a diffusion neural network and use it to g...

arxiv.org/abs/2406.04662v1

Evaluating and Mitigating IP Infringement in Visual Generative AI

The popularity of visual generative AI models like DALL-E 3, Stable Diffusion XL, Stable Video Diffusion, and Sora has been increasing. Through extensive evaluation, we discovered that the state-of-the-art visual generative models can generate conten...