1,455 results for diffusion

github.com/PaulCouairon/DiffCut

PaulCouairon/DiffCut

[NeurIPS 2024] Official code for DiffCut: Catalyzing Zero-Shot Semantic Segmentation with Diffusion Features and Recursive Normalized Cut (⭐ 50)

arxiv.org/abs/2404.14700v4

FlashSpeech: Efficient Zero-Shot Speech Synthesis

Recent progress in large-scale zero-shot speech synthesis has been significantly advanced by language models and diffusion models. However, the generation process of both methods is slow and computationally intensive. Efficient speech synthesis using...

arxiv.org/abs/2507.23620v1

DivControl: Knowledge Diversion for Controllable Image Generation

Diffusion models have advanced from text-to-image (T2I) to image-to-image (I2I) generation by incorporating structured inputs such as depth maps, enabling fine-grained spatial control. However, existing methods either train separate models for each c...

arxiv.org/abs/1703.08593v2

Analyzing Evolving Stories in News Articles

There is an overwhelming number of news articles published every day around the globe. Following the evolution of a news-story is a difficult task given that there is no such mechanism available to track back in time to study the diffusion of the rel...

arxiv.org/abs/2502.18477v3

Personalized Image Generation for Recommendations Beyond Catalogs

Personalization is central to human-AI interaction, yet current diffusion-based image generation systems remain largely insensitive to user diversity. Existing attempts to address this often rely on costly paired preference data or introduce latency...

arxiv.org/abs/2209.14746v1

Diffusion-assisted molecular beam epitaxy of CuCrO$_2$ thin films

Using molecular beam epitaxy (MBE) to grow multi-elemental oxides (MEO) is generally challenging, partly due to difficulty in stoichiometry control. Occasionally, if one of the elements is volatile at the growth temperature, stoichiometry control can...

arxiv.org/abs/2309.03335v2

SADIR: Shape-Aware Diffusion Models for 3D Image Reconstruction

3D image reconstruction from a limited number of 2D images has been a long-standing challenge in computer vision and image analysis. While deep learning-based approaches have achieved impressive performance in this area, existing deep networks often...

arxiv.org/abs/2403.16111v1

EVA: Zero-shot Accurate Attributes and Multi-Object Video Editing

Current diffusion-based video editing primarily focuses on local editing (\textit{e.g.,} object/background editing) or global style editing by utilizing various dense correspondences. However, these methods often fail to accurately edit the foregroun...

arxiv.org/abs/2509.23165v2

Untangling Vascular Trees for Surgery and Interventional Radiology

The diffusion of minimally invasive, endovascular interventions motivates the development of visualization methods for complex vascular networks. We propose a planar representation of blood vessel trees which preserves the properties that are most re...

arxiv.org/abs/2107.08866v2

Quantum walk on a comb with infinite teeth

We study continuous time quantum random walk on a comb with infinite teeth and show that the return probability to the starting point decays with time $t$ as $t^{-1}$. We analyse the diffusion along the spine and into the teeth and show that the walk...