arxiv.org/abs/2009.12165v1
During the winter season, real-time monitoring of road surface conditions is critical for the safety of drivers and road maintenance operations. Previous research has evaluated the potential of image classification methods for detecting road snow cov...
github.com/nchamo/DroneMissionSimulator
Simulate drone flights and take images of your 3D objects (⭐ 6)
arxiv.org/abs/2106.05058v1
With the recent surge in the research of vision transformers, they have demonstrated remarkable potential for various challenging computer vision applications, such as image recognition, point cloud classification as well as video understanding. In t...
arxiv.org/abs/1912.11658v1
Nowadays document analysis and recognition remain challenging tasks. However, only a few datasets designed for text detection (TD) and optical character recognition (OCR) problems exist. In this paper we present Distorted Document Images dataset (DDI...
www.k2.t.u-tokyo.ac.jp/vision/DPM/
Points: 1306 | Comments: 128 | Author: hongzi
arxiv.org/abs/2212.07664v1
The analysis of digitized historical manuscripts is typically addressed by paleographic experts. Writer identification refers to the classification of known writers while writer retrieval seeks to find the writer by means of image similarity in a dat...
arxiv.org/abs/2202.05331v2
Several services for people with visual disabilities have emerged recently due to achievements in Assistive Technologies and Artificial Intelligence areas. Despite the growth in assistive systems availability, there is a lack of services that support...
arxiv.org/abs/2510.08442v2
Visual Reinforcement Learning (RL) agents must learn to act based on high-dimensional image data where only a small fraction of the pixels is task-relevant. This forces agents to waste exploration and computational resources on irrelevant features, l...
arxiv.org/abs/2401.09773v2
Nuclei instance segmentation in histopathological images is of great importance for biological analysis and cancer diagnosis but remains challenging for two reasons. (1) Similar visual presentation of intranuclear and extranuclear regions of chromoph...
www.bing.com/ck/a?!&&p=70d954baf5ca8dff0c7d04dbe0c0bf993b33b677112cf4c251e5eb00e2e1c7a1JmltdHM9MTc3MjQ5NjAwMA&ptn=3&ver=2&hsh=4&fclid=09179e2a-c0fd-66d8-22a7-8938c1986796&u=a1aHR0cHM6Ly9lamplLndlYmxpby5qcC9jb250ZW50L2Vudmlyb25tZW50K2NvbXBhdGlibGU&ntb=1
To simultaneously attain high image quality processing and time reduction from application of power which is compatible with an environment to start of operation.
www.bing.com/ck/a?!&&p=039c9fc703a36028b9e734f93630355de65ece885cde44994aab2d068e7a21dbJmltdHM9MTc3MjQ5NjAwMA&ptn=3&ver=2&hsh=4&fclid=06733659-74e9-6a6b-20c5-214b75996bda&u=a1aHR0cHM6Ly93d3cuYmluZy5jb20vbWFwcy9kaXJlY3Rpb25z&ntb=1
Map multiple locations, get transit/walking/driving directions, view live traffic conditions, plan trips, view satellite, aerial and 3d imagery. Do more with Bing Maps.
github.com/bharat-b7/MultiGarmentNetwork
Repo for "Multi-Garment Net: Learning to Dress 3D People from Images, ICCV'19" (⭐ 304)
arxiv.org/abs/1907.13615v3
Three-dimensional human body models are widely used in the analysis of human pose and motion. Existing models, however, are learned from minimally-clothed 3D scans and thus do not generalize to the complexity of dressed people in common images and vi...
arxiv.org/abs/1203.5347v1
The focus of this work is to demonstrate how spatially resolved image information from diesel fuel injection events can be obtained using a forward-scatter imaging geometry, and used to calculate the velocities of liquid structures on the periphery o...
github.com/s-huu/TurningWeaknessIntoStrength
Official implementation for paper: A New Defense Against Adversarial Images: Turning a Weakness into a Strength (⭐ 38)
www.bbc.com/news/in-pictures-56238018
Points: 1032 | Comments: 347 | Author: astdb
arxiv.org/abs/2504.01955v1
Unsupervised panoptic segmentation aims to partition an image into semantically meaningful regions and distinct object instances without training on manually annotated data. In contrast to prior work on unsupervised panoptic scene understanding, we e...
arxiv.org/abs/2507.06230v2
Semantic scene completion (SSC) aims to infer both the 3D geometry and semantics of a scene from single images. In contrast to prior work on SSC that heavily relies on expensive ground-truth annotations, we approach SSC in an unsupervised setting. Ou...
arxiv.org/abs/2106.08285v3
Time-lapse fluorescent microscopy (TLFM) combined with predictive mathematical modelling is a powerful tool to study the inherently dynamic processes of life on the single-cell level. Such experiments are costly, complex and labour intensive. A compl...
arxiv.org/abs/2304.07597v4
Extracting single-cell information from microscopy data requires accurate instance-wise segmentations. Obtaining pixel-wise segmentations from microscopy imagery remains a challenging task, especially with the added complexity of microstructured envi...