arxiv.org/abs/2509.08422v4
Video-based AI systems are increasingly adopted in safety-critical domains such as autonomous driving and healthcare. However, interpreting their decisions remains challenging due to the inherent spatiotemporal complexity of video data and the opacit...
arxiv.org/abs/2512.15176v1
Efficiency, as a critical practical challenge for LLM-driven agentic and reasoning systems, is increasingly constrained by the inherent latency of autoregressive (AR) decoding. Speculative decoding mitigates this cost through a draft-verify scheme, y...
arxiv.org/abs/1204.3230v1
The rapid diffusion of "microblogging" services such as Twitter is ushering in a new era of possibilities for organizations to communicate with and engage their core stakeholders and the general public. To enhance understanding of the communicative f...
github.com/JIA-Lab-research/RIVAL
[NeurIPS 2023 Spotlight] Real-World Image Variation by Aligning Diffusion Inversion Chain (⭐ 153)
arxiv.org/abs/2503.12615v2
Text-to-image latent diffusion models (LDMs) have recently emerged as powerful generative models with great potential for solving inverse problems in imaging. However, leveraging such models in a Plug & Play (PnP), zero-shot manner remains challengin...
arxiv.org/abs/2501.13920v1
With the rapid development of diffusion models, text-to-image(T2I) models have made significant progress, showcasing impressive abilities in prompt following and image generation. Recently launched models such as FLUX.1 and Ideogram2.0, along with ot...
arxiv.org/abs/2009.10046v1
We propose a model for glioma patterns in a microlocal tumor environment under the influence of acidity, angiogenesis, and tissue anisotropy. The bottom-up model deduction eventually leads to a system of reaction-diffusion-taxis equations for glioma...
arxiv.org/abs/2211.15271v2
The paper discusses the potential of large vision-language models as objects of interest for empirical cultural studies. Focusing on the comparative analysis of outputs from two popular text-to-image synthesis models, DALL-E 2 and Stable Diffusion, t...
arxiv.org/abs/2507.23001v1
Deep learning models for skin disease classification require large, diverse, and well-annotated datasets. However, such resources are often limited due to privacy concerns, high annotation costs, and insufficient demographic representation. While tex...
arxiv.org/abs/1802.05149v1
We present an agent-based model to simulate gang territorial development motivated by graffiti marking on a two-dimensional discrete lattice. For simplicity, we assume that there are two rival gangs present, and they compete for territory. In this mo...
www.bing.com/ck/a?!&&p=0b5df14c296014ad54ee4454b492dc732fc8b455c848f1afbd37cb25f36ef3dbJmltdHM9MTc3MjQ5NjAwMA&ptn=3&ver=2&hsh=4&fclid=20800594-f661-692e-3f78-1285f7326895&u=a1aHR0cHM6Ly9uZXdzLm1pdC5lZHUvMjAyNS9haS10b29sLWdlbmVyYXRlcy1oaWdoLXF1YWxpdHktaW1hZ2VzLWZhc3Rlci0wMzIx&ntb=1
Mar 21, 2025 · A hybrid AI approach known as hybrid autoregressive transformer can generate realistic images with the same or better quality than state-of-the-art diffusion models, but that runs about nine …
www.bing.com/ck/a?!&&p=40dabe6091f6a8712fd6eab4d9cfd995801ce4b603519ea07205ae9bcb0af86cJmltdHM9MTc3MjQ5NjAwMA&ptn=3&ver=2&hsh=4&fclid=20800594-f661-692e-3f78-1285f7326895&u=a1aHR0cHM6Ly9uZXdzLm1pdC5lZHUvdG9waWMvYXJ0aWZpY2lhbC1pbnRlbGxpZ2VuY2Uy&ntb=1
4 days ago · AI algorithm enables tracking of vital white matter pathways Opening a new window on the brainstem, a new tool reliably and finely resolves distinct nerve bundles in live diffusion MRI scans, …
arxiv.org/abs/2602.19706v1
Single LDR to HDR reconstruction remains challenging for over-exposed regions where traditional methods often fail due to complete information loss. We present a training-free approach that enhances existing indirect and direct HDR reconstruction met...
github.com/facebookresearch/AGRoL
Code release for "Avatars Grow Legs Generating Smooth Human Motion from Sparse Tracking Inputs with Diffusion Model", CVPR 2023 (⭐ 250)
arxiv.org/abs/1407.7694v2
Vacancy diffusion and clustering processes in body-centered-cubic (bcc) Fe are studied using the kinetic activation-relaxation technique (k-ART), an off-lattice kinetic Monte Carlo method with on-the-fly catalog building capabilities. For monovacanci...
github.com/km1994/LLMsNineStoryDemonTower
【LLMs九层妖塔】分享 LLMs在自然语言处理(ChatGLM、Chinese-LLaMA-Alpaca、小羊驼 Vicuna、LLaMA、GPT4ALL等)、信息检索(langchain)、语言合成、语言识别、多模态等领域(Stable Diffusion、MiniGPT-4、VisualGLM-6B、Ziya-Visual等)等 实战与经验。 (⭐ 2156)
arxiv.org/abs/1703.04420v1
We perform mathematical anaysis of the biofilm development process. A model describing biomass growth is proposed: It arises from coupling three parabolic nonlinear equations: a biomass equation with degenerate and singular diffusion, a nutrient tran...
arxiv.org/abs/2410.05027v1
Automatic magnetic resonance (MR) image processing pipelines are widely used to study people with multiple sclerosis (PwMS), encompassing tasks such as lesion segmentation and brain parcellation. However, the presence of lesion often complicates thes...
arxiv.org/abs/2503.24210v1
Reconstructing sharp 3D representations from blurry multi-view images are long-standing problem in computer vision. Recent works attempt to enhance high-quality novel view synthesis from the motion blur by leveraging event-based cameras, benefiting f...
arxiv.org/abs/2510.00506v1
How can we reconstruct 3D hand poses when large portions of the hand are heavily occluded by itself or by objects? Humans often resolve such ambiguities by leveraging contextual knowledge -- such as affordances, where an object's shape and function s...