3,157 results for vision · 0.158s

arxiv.org/abs/2309.01617v1

DeViL: Decoding Vision features into Language

Post-hoc explanation methods have often been criticised for abstracting away the decision-making process of deep neural networks. In this work, we would like to provide natural language descriptions for what different layers of a vision backbone have...

Sponsored Partners
arxiv.org/abs/1508.00102v1

Towards Distortion-Predictable Embedding of Neural Networks

Current research in Computer Vision has shown that Convolutional Neural Networks (CNN) give state-of-the-art performance in many classification tasks and Computer Vision problems. The embedding of CNN, which is the internal representation produced by...

github.com/landing-ai/vision-agent

landing-ai/vision-agent

This tool has been deprecated. Use Agentic Document Extraction instead. (⭐ 5262)

www.bing.com/ck/a?!&&p=0b04708bfb5a8b8d48c22cc65020d22c56d815c00f8107d2a1ec56f2147e0300JmltdHM9MTc3MjY2ODgwMA&ptn=3&ver=2&hsh=4&fclid=0a90191a-4486-6b47-05fc-0e0e45e26a12&u=a1aHR0cHM6Ly93d3cuYXBwbGUuY29tL3Nob3AvdmlzaW9uL2FjY2Vzc29yaWVz&ntb=1

Buy Apple Vision Pro Accessories

Explore accessories to get the most out of Apple Vision Pro. Shop cables, gaming controllers, and more. Get fast, free shipping at apple.com.

www.bing.com/ck/a?!&&p=b4a648182a8b056e5d65cce19e5ee26ac1b49accfe6499570cba45df08b3c955JmltdHM9MTc3MjY2ODgwMA&ptn=3&ver=2&hsh=4&fclid=0a90191a-4486-6b47-05fc-0e0e45e26a12&u=a1aHR0cHM6Ly93d3cuYXBwbGUuY29tL3Nob3AvYWNjZXNzb3JpZXMvYWxs&ntb=1

Apple Accessories for Apple Watch, iPhone, iPad, Mac and Vision Pro

Shop Apple accessories for Apple Watch, iPhone, iPad, Mac, and Vision Pro. Search by product lines or browse by categories. Buy now with fast, free shipping.

github.com/computer-vision/takahashi2012cvpr

computer-vision/takahashi2012cvpr

⚠️ This repository is archived. Please use https://github.com/nbhr/pycalib . An Implementation of Takahashi, Nobuhara and Matsuyama "A New Mirror-based Camera Pose Estimation Using an Orthogonality Constraint" presented at CVPR 2012 (⭐ 33)

arxiv.org/abs/2404.09933v1

HOI-Ref: Hand-Object Interaction Referral in Egocentric Vision

Large Vision Language Models (VLMs) are now the de facto state-of-the-art for a number of tasks including visual question answering, recognising objects, and spatial referral. In this work, we propose the HOI-Ref task for egocentric images that aims...

arxiv.org/abs/2508.12268v2

iTrace: Click-Based Gaze Visualization on the Apple Vision Pro

The Apple Vision Pro is equipped with accurate eye-tracking capabilities, yet the privacy restrictions on the device prevent direct access to continuous user gaze data. This study introduces iTrace, a novel application that overcomes these limitation...

arxiv.org/abs/2510.18034v1

SAVANT: Semantic Analysis with Vision-Augmented Anomaly deTection

Autonomous driving systems remain critically vulnerable to the long-tail of rare, out-of-distribution scenarios with semantic anomalies. While Vision Language Models (VLMs) offer promising reasoning capabilities, naive prompting approaches yield unre...

github.com/arcVaishali/stellar-vision

arcVaishali/stellar-vision

This is a disaster relief and rescue aid. It aims to reduce the issue of fragmented response by various organizations during the times of natural calamity. IMPACT HACKS'23 First Runner up (⭐ 6)

arxiv.org/abs/2502.17092v1

Shakti-VLMs: Scalable Vision-Language Models for Enterprise AI

We introduce Shakti VLM, a family of vision-language models in the capacity of 1B and 4B parameters designed to address data efficiency challenges in multimodal learning. While recent VLMs achieve strong performance through extensive training data, S...