1,455 results for diffusion

github.com/Will-Consagra/OpEdd

Will-Consagra/OpEdd

Optimal experimental design and sparse estimation for diffusion MRI (⭐ 0)

github.com/IDEA-Research/Grounded-Segment-Anything

IDEA-Research/Grounded-Segment-Anything

Grounded SAM: Marrying Grounding DINO with Segment Anything & Stable Diffusion & Recognize Anything - Automatically Detect , Segment and Generate Anything (⭐ 17446)

arxiv.org/abs/2007.06752v1

Speeding Up Particle Slowing using Shortcuts to Adiabaticity

We propose a method for slowing particles by laser fields that potentially has the ability to generate large forces without the associated momentum diffusion that results from the random directions of spontaneously scattered photons. In this method,...

arxiv.org/abs/1508.03225v1

A non-smooth regularization of a forward-backward parabolic equation

In this paper we introduce a model describing diffusion of species by a suitable regularization of a "forward-backward" parabolic equation. In particular, we prove existence and uniqueness of solutions, as well as continuous dependence on data, for a...

github.com/yisol/IDM-VTON

yisol/IDM-VTON

[ECCV2024] IDM-VTON : Improving Diffusion Models for Authentic Virtual Try-on in the Wild (⭐ 4898)

arxiv.org/abs/2505.05022v2

SOAP: Style-Omniscient Animatable Portraits

Creating animatable 3D avatars from a single image remains challenging due to style limitations (realistic, cartoon, anime) and difficulties in handling accessories or hairstyles. While 3D diffusion models advance single-view reconstruction for gener...

arxiv.org/abs/2405.19335v1

X-VILA: Cross-Modality Alignment for Large Language Model

We introduce X-VILA, an omni-modality model designed to extend the capabilities of large language models (LLMs) by incorporating image, video, and audio modalities. By aligning modality-specific encoders with LLM inputs and diffusion decoders with LL...

arxiv.org/abs/1312.0650v1

Differential Games of Competition in Online Content Diffusion

Access to online contents represents a large share of the Internet traffic. Most such contents are multimedia items which are user-generated, i.e., posted online by the contents' owners. In this paper we focus on how those who provide contents can le...

github.com/design-edit/DesignEdit

design-edit/DesignEdit

[AAAI2025] DesignEdit: Unify Spatial-Aware Image Editing via Training-free Inpainting with a Multi-Layered Latent Diffusion Framework (⭐ 365)

github.com/nv-tlabs/pacer

nv-tlabs/pacer

Official implementation of PACER, Pedestrian Animation ControllER, of CVPR 2023 paper: "Trace and Pace: Controllable Pedestrian Animation via Guided Trajectory Diffusion". (⭐ 107)