3,157 results for vision · 0.158s

arxiv.org/abs/2411.06657v2

Renaissance: Investigating the Pretraining of Vision-Language Encoders

In the past several years there has been an explosion of available models for vision-language (VL) tasks. Unfortunately, the literature still leaves open a number of questions related to best practices in designing and training such models. Additiona...

arxiv.org/abs/2309.00035v1

FACET: Fairness in Computer Vision Evaluation Benchmark

Computer vision models have known performance disparities across attributes such as gender and skin tone. This means during tasks such as classification and detection, model performance differs for certain classes based on the demographics of the peo...

Sponsored Partners
arxiv.org/abs/2309.11751v2

How Robust is Google's Bard to Adversarial Image Attacks?

Multimodal Large Language Models (MLLMs) that integrate text and other modalities (especially vision) have achieved unprecedented performance in various multimodal tasks. However, due to the unsolved adversarial robustness problem of vision models, M...

github.com/visionmedia/page.js

visionmedia/page.js

Micro client-side router inspired by the Express router (⭐ 7698)

arxiv.org/abs/2112.10381v1

Automated Vision-Based Wellness Analysis for Elderly Care Centers

The growth in the aging population requires caregivers to improve both efficiency and quality of healthcare. In this study, we develop an automatic, vision-based system for monitoring and analyzing the physical and mental well-being of senior citizen...

arxiv.org/abs/2507.01654v1

SPoT: Subpixel Placement of Tokens in Vision Transformers

Vision Transformers naturally accommodate sparsity, yet standard tokenization methods confine features to discrete patch grids. This constraint prevents models from fully exploiting sparse regimes, forcing awkward compromises. We propose Subpixel Pla...

github.com/Tunnel-vision/sinaclassification_scrapyspdier

Tunnel-vision/sinaclassification_scrapyspdier

新浪网分类资讯爬虫 爬取新浪网导航页所有下所有大类、小类、小类里的子链接,以及子链接页面的新闻内容。 (⭐ 2)

www.bing.com/ck/a?!&&p=d316ec0e261876c7303070e3307b9ed473057e2ac94a6e81e2e5f9b9706a2ddeJmltdHM9MTc3Mjg0MTYwMA&ptn=3&ver=2&hsh=4&fclid=153d806e-aa16-6f5b-2469-977babeb6e21&u=a1aHR0cHM6Ly93d3cuYmlibGVnYXRld2F5LmNvbS9wYXNzYWdlLz9zZWFyY2g9SXNhaWFoJTIwMSZ2ZXJzaW9uPU5JVg&ntb=1

Isaiah 1 NIV - The vision concerning Judah and - Bible Gateway

The vision concerning Judah and Jerusalem that Isaiah son of Amoz saw during the reigns of Uzziah, Jotham, Ahaz and Hezekiah, kings of Judah. A

arxiv.org/abs/2408.10388v1

Narrowing the Gap between Vision and Action in Navigation

The existing methods for Vision and Language Navigation in the Continuous Environment (VLN-CE) commonly incorporate a waypoint predictor to discretize the environment. This simplifies the navigation actions into a view selection task and improves nav...

www.bing.com/ck/a?!&&p=8a5a403f04e75688164ba7fd7e44ee14903f94cee2b67106aa9492d47b9e04a2JmltdHM9MTc3Mjg0MTYwMA&ptn=3&ver=2&hsh=4&fclid=33bfc0f6-096f-6b2c-0cbe-d7e3089b6a35&u=a1aHR0cHM6Ly93d3cuYmlibGVnYXRld2F5LmNvbS9wYXNzYWdlLz9zZWFyY2g9SXNhaWFoJTIwMSZ2ZXJzaW9uPU5JVg&ntb=1

Isaiah 1 NIV - The vision concerning Judah and - Bible Gateway

The vision concerning Judah and Jerusalem that Isaiah son of Amoz saw during the reigns of Uzziah, Jotham, Ahaz and Hezekiah, kings of Judah. A

arxiv.org/abs/2505.15441v4

Octic Vision Transformers: Quicker ViTs Through Equivariance

Why are state-of-the-art Vision Transformers (ViTs) not designed to exploit natural geometric symmetries such as 90-degree rotations and reflections? In this paper, we argue that there is no fundamental reason, and what has been missing is an efficie...

en.wikipedia.org/wiki/Converge_Vision

Converge Vision - Wikipedia

Converge Vision, separately marketed as Sky TV (for Metro Manila) and Converge FiberTV (for select areas in key provinces), is a digital Internet Protocol