arxiv.org/abs/1405.5047v2
Traditional approaches to upper body pose estimation using monocular vision rely on complex body models and a large variety of geometric constraints. We argue that this is not ideal and somewhat inelegant as it results in large processing burdens, an...
arxiv.org/abs/2501.17585v2
This paper presents the design and implementation of TAPOR, a privacy-preserving, non-contact, and fully passive sensing system for accurate and robust 3D hand pose reconstruction for around-device interaction using a single low-cost thermal array se...
arxiv.org/abs/1803.08225v1
We present a box-free bottom-up approach for the tasks of pose estimation and instance segmentation of people in multi-person images using an efficient single-shot model. The proposed PersonLab model tackles both semantic-level reasoning and object-p...
arxiv.org/abs/2305.03209v3
We consider the 2d $β$-plane stochastic Navier-Stokes equations in a periodic channel. We prove the well-posedness and existence of the stationary measure, as well as certain regularity estimates concerning the support of the stationary measure. The...
arxiv.org/abs/2406.04754v2
This paper studies the global well-posedness and optimal decay estimates to the Oldroyd-B model in $\mathbb R^d$ ($d\geq2$). By utilizing the special structure of this system, we give a simplified proof to the global existence of solutions for the ca...
arxiv.org/abs/2005.11780v1
Head pose estimation is a crucial problem for many tasks, such as driver attention, fatigue detection, and human behaviour analysis. It is well known that neural networks are better at handling classification problems than regression problems. It is...
arxiv.org/abs/1906.07251v2
Generating a photorealistic image with intended human pose is a promising yet challenging research topic for many applications such as smart photo editing, movie making, virtual try-on, and fashion display. In this paper, we present a novel deep gene...
arxiv.org/abs/2507.08396v2
Subject-consistent generation (SCG)-aiming to maintain a consistent subject identity across diverse scenes-remains a challenge for text-to-image (T2I) models. Existing training-free SCG methods often achieve consistency at the cost of layout and pose...
arxiv.org/abs/1707.02439v2
This paper presents a deep learning based approach to the problem of human pose estimation. We employ generative adversarial networks as our learning paradigm in which we set up two stacked hourglass networks with the same architecture, one as the ge...
arxiv.org/abs/1507.00302v1
We present a method for learning an embedding that places images of humans in similar poses nearby. This embedding can be used as a direct method of comparing images based on human pose, avoiding potential challenges of estimating body joint position...
arxiv.org/abs/2507.22742v1
Accurate human trajectory prediction is one of the most crucial tasks for autonomous driving, ensuring its safety. Yet, existing models often fail to fully leverage the visual cues that humans subconsciously communicate when navigating the space. In...
arxiv.org/abs/2104.11116v1
While accurate lip synchronization has been achieved for arbitrary-subject audio-driven talking face generation, the problem of how to efficiently drive the head pose remains. Previous methods rely on pre-estimated structural information such as land...
arxiv.org/abs/2502.02181v1
We prove well-posedness for higher-order equations in the so-called dNLS hierarchy (also known as part of the Kaup-Newell hierarchy) in almost critical Fourier-Lebesgue and in modulation spaces. Leaning in on estimates proven by the author in a previ...
arxiv.org/abs/1501.04540v2
For any graded poset $P$, we define a new graded poset, $\mathcal E(P)$, whose elements are the edges in the Hasse diagram of P. For any group, $G$, acting on the boolean algebra, $B_n$, we conjecture that $\mathcal E(B_n/G)$ is Peck. We prove that t...
arxiv.org/abs/2412.06029v1
Precise camera pose control is crucial for video generation with diffusion models. Existing methods require fine-tuning with additional datasets containing paired videos and camera pose annotations, which are both data-intensive and computationally c...
arxiv.org/abs/2404.02544v4
Existing research on unconstrained in-the-wild head pose estimation suffers from the flaws of its datasets, which consist of either numerous samples by non-realistic synthesis or constrained collection, or small-scale natural images yet with plausibl...
arxiv.org/abs/2003.02987v2
The hard phase model describes a relativistic barotropic irrotational fluid with sound speed equal to the speed of light. In this paper, we prove the local well-posedness for this model in the Minkowski background with free boundary. Moreover, we sho...
arxiv.org/abs/2104.11931v1
We propose an approach to generate images of people given a desired appearance and pose. Disentangled representations of pose and appearance are necessary to handle the compound variability in the resulting generated images. Hence, we develop an appr...
arxiv.org/abs/2411.00833v1
Yoga has recently become an essential aspect of human existence for maintaining a healthy body and mind. People find it tough to devote time to the gym for workouts as their lives get more hectic and they work from home. This kind of human pose estim...
arxiv.org/abs/2306.15768v1
Pose recognition deals with designing algorithms to locate human body joints in a 2D/3D space and run inference on the estimated joint locations for predicting the poses. Yoga poses consist of some very complex postures. It imposes various challenges...