1,863 results for Recognition · 0.121s

Sponsored Partners
arxiv.org/abs/1912.11658v1

DDI-100: Dataset for Text Detection and Recognition

Nowadays document analysis and recognition remain challenging tasks. However, only a few datasets designed for text detection (TD) and optical character recognition (OCR) problems exist. In this paper we present Distorted Document Images dataset (DDI...

github.com/zzw922cn/awesome-speech-recognition-speech-synthesis-papers

zzw922cn/awesome-speech-recognition-speech-synthesis-papers

Automatic Speech Recognition (ASR), Speaker Verification, Speech Synthesis, Text-to-Speech (TTS), Language Modelling, Singing Voice Synthesis (SVS), Voice Conversion (VC) (⭐ 3118)

en.wikipedia.org/wiki/Recognition-primed_decision

Recognition-primed decision - Wikipedia

Recognition-primed decision (RPD) is a model of how people make quick, effective decisions when faced with complex situations. In this model, the decision

arxiv.org/abs/1304.7184v1

Reading Ancient Coin Legends: Object Recognition vs. OCR

Standard OCR is a well-researched topic of computer vision and can be considered solved for machine-printed text. However, when applied to unconstrained images, the recognition rates drop drastically. Therefore, the employment of object recognition-b...

arxiv.org/abs/2102.04960v2

Radar-to-Lidar: Heterogeneous Place Recognition via Joint Learning

Place recognition is critical for both offline mapping and online localization. However, current single-sensor based place recognition still remains challenging in adverse conditions. In this paper, a heterogeneous measurements based framework is pro...

www.bing.com/ck/a?!&&p=7cf221e6512def2468775746cf7583e2bbcdd2b6f51533cd2b911568e319f286JmltdHM9MTc3MjQ5NjAwMA&ptn=3&ver=2&hsh=4&fclid=3cf07a33-1a82-6f9e-29d4-6d221b0c6eed&u=a1aHR0cHM6Ly9naXRodWIuY29tL29wZW5haS93aGlzcGVy&ntb=1

GitHub - openai/whisper: Robust Speech Recognition via Large-Scale …

Whisper is a general-purpose speech recognition model. It is trained on a large dataset of diverse audio and is also a multitasking model that can perform multilingual speech recognition, speech translation, …

arxiv.org/abs/2302.03873v1

Geometric Perception based Efficient Text Recognition

Every Scene Text Recognition (STR) task consists of text localization \& text recognition as the prominent sub-tasks. However, in real-world applications with fixed camera positions such as equipment monitor reading, image-based data entry, and print...

arxiv.org/abs/1610.04575v1

Comparing Face Detection and Recognition Techniques

This paper implements and compares different techniques for face detection and recognition. One is find where the face is located in the images that is face detection and second is face recognition that is identifying the person. We study three techn...