arxiv.org/abs/2501.05442v2
Video tokenizers are essential for latent video diffusion models, converting raw video data into spatiotemporally compressed latent spaces for efficient training. However, extending state-of-the-art video tokenizers to achieve a temporal compression...
www.bing.com/ck/a?!&&p=2398743af36f6d0950f272694a49016e69e024fa19a24ce179f8eb7e49be5151JmltdHM9MTc3MjU4MjQwMA&ptn=3&ver=2&hsh=4&fclid=29a908c5-bd91-6a15-236b-1fd7bc376b1c&u=a1aHR0cHM6Ly9uZXdzLm1pdC5lZHUvMjAyNS9haS10b29sLWdlbmVyYXRlcy1oaWdoLXF1YWxpdHktaW1hZ2VzLWZhc3Rlci0wMzIx&ntb=1
Mar 21, 2025 · A hybrid AI approach known as hybrid autoregressive transformer can generate realistic images with the same or better quality than state-of-the-art diffusion models, but that runs about nine …
www.bing.com/ck/a?!&&p=53c425fca550ea7a6edb80cf2d9a717aa4a9fd8cee1687ff366dc775b7b894adJmltdHM9MTc3MjU4MjQwMA&ptn=3&ver=2&hsh=4&fclid=29a908c5-bd91-6a15-236b-1fd7bc376b1c&u=a1aHR0cHM6Ly9uZXdzLm1pdC5lZHUvdG9waWMvYXJ0aWZpY2lhbC1pbnRlbGxpZ2VuY2Uy&ntb=1
5 days ago · AI algorithm enables tracking of vital white matter pathways Opening a new window on the brainstem, a new tool reliably and finely resolves distinct nerve bundles in live diffusion MRI scans, …
arxiv.org/abs/2501.07533v1
Canine cardiomegaly, marked by an enlarged heart, poses serious health risks if undetected, requiring accurate diagnostic methods. Current detection models often rely on small, poorly annotated datasets and struggle to generalize across diverse imagi...
www.bing.com/ck/a?!&&p=9ac950c03593a2c265028564fb532662ef5954330158ac6c44926df7ba4a3f92JmltdHM9MTc3MjU4MjQwMA&ptn=3&ver=2&hsh=4&fclid=25ecfac8-d03a-6663-1c46-eddad1ec67e9&u=a1aHR0cHM6Ly93d3cuY29udGVudHMuY29tL2l0L2VudGVycHJpc2Uv&ntb=1
Con Contents, valorizzi il lavoro di creazione dei contenuti mantenendo l'integrità del marchio e potenziando la sua diffusione globale, attraverso una piattaforma unica e centralizzata.
arxiv.org/abs/2507.17909v1
We consider a domain $Ω\subseteq\mathbb{R\!}^{\,2}$ with branched fractal boundary $Γ^{\infty}$ and parameter $τ\in[1/2,τ^{\ast}]$ introduced by Achdou and Tchou \cite{ACH08}, for $τ^{\ast}\simeq 0.593465$, which acts as an idealization of the b...
arxiv.org/abs/1205.4979v2
Existence and uniqueness are investigated for a nonlinear diffusion problem of phase-field type, consisting of a parabolic system of two partial differential equations, complemented by Neumann homogeneous boundary conditions and initial conditions. T...
arxiv.org/abs/0909.1055v1
Equation for anomalous diffusion in momentum space, recently obtained in the recent paper (S.A. Trigger, ArXiv 0907.2793 v1, [cond-matt. stat.-mech.], 16 July 2009) is solved for the stationary and non-stationary cases on basis of the appropriate p...
arxiv.org/abs/2509.18433v1
Utilizing offline reinforcement learning (RL) with real-world clinical data is getting increasing attention in AI for healthcare. However, implementation poses significant challenges. Defining direct rewards is difficult, and inverse RL (IRL) struggl...
github.com/ai-forever/KandinskyVideo
KandinskyVideo — multilingual end-to-end text2video latent diffusion model (⭐ 186)
github.com/kandinskylab/kandinsky-5
Kandinsky 5.0: A family of diffusion models for Video & Image generation (⭐ 718)
github.com/ai-forever/Kandinsky-2
Kandinsky 2 — multilingual text2image latent diffusion model (⭐ 2819)
www.bing.com/ck/a?!&&p=675df7c129acd8413958fc5bef48eaef367a5d8b39b59a412dd501daca588410JmltdHM9MTc3MjU4MjQwMA&ptn=3&ver=2&hsh=4&fclid=3d40134a-74e0-6296-0afe-04587554636c&u=a1aHR0cHM6Ly93d3cuY29udGVudHMuY29tL2l0L2VudGVycHJpc2Uv&ntb=1
Con Contents, valorizzi il lavoro di creazione dei contenuti mantenendo l'integrità del marchio e potenziando la sua diffusione globale, attraverso una piattaforma unica e centralizzata.
arxiv.org/abs/2403.05094v1
Face personalization aims to insert specific faces, taken from images, into pretrained text-to-image diffusion models. However, it is still challenging for previous methods to preserve both the identity similarity and editability due to overfitting t...
arxiv.org/abs/2105.14696v1
Subject of this work is to investigate the kinetics of mass transfer of volatile amphiphiles from their vapors to aqueous drops, and from the saturated aqueous drop solutions to air. The used amphiphiles are benzyl acetate, linalool, and citronellol....
arxiv.org/abs/2003.01593v1
Artificial intelligence shows promise for solving many practical societal problems in areas such as healthcare and transportation. However, the current mechanisms for AI model diffusion such as Github code repositories, academic project webpages, and...
arxiv.org/abs/2012.12004v2
We employ the epidemic Renormalization Group (eRG) framework to understand, reproduce and predict the COVID-19 pandemic diffusion across the US. The human mobility across different geographical US divisions is modelled via open source flight data alo...
arxiv.org/abs/2510.10135v2
Ensuring character identity consistency across varying prompts remains a fundamental limitation in diffusion-based text-to-image generation. We propose CharCom, a modular and parameter-efficient framework that achieves character-consistent story illu...
arxiv.org/abs/2503.15996v1
Animation of humanoid characters is essential in various graphics applications, but requires significant time and cost to create realistic animations. We propose an approach to synthesize 4D animated sequences of input static 3D humanoid meshes, leve...
arxiv.org/abs/2409.09149v1
Diffusion models have shown their remarkable ability to synthesize images, including the generation of humans in specific poses. However, current models face challenges in adequately expressing conditional control for detailed hand pose generation, l...