860 results for Evaluating · 0.092s

Sponsored Partners
arxiv.org/abs/1709.10147v1

Maat: A Platform Service for Measurement and Attestation

Software integrity measurement and attestation (M&A) are critical technologies for evaluating the trustworthiness of software platforms. To best support these technologies, next generation systems must provide a centralized service for securely selec...

arxiv.org/abs/2506.23963v1

Evaluating the Impact of Khmer Font Types on Text Recognition

Text recognition is significantly influenced by font types, especially for complex scripts like Khmer. The variety of Khmer fonts, each with its unique character structure, presents challenges for optical character recognition (OCR) systems. In this...

github.com/thu-coai/Safety-Prompts

thu-coai/Safety-Prompts

Chinese safety prompts for evaluating and improving the safety of LLMs. 中文安全prompts,用于评估和提升大模型的安全性。 (⭐ 1132)

www.bing.com/ck/a?!&&p=84f145607e9549f079c8baa284114849892ba25f03c2993814f15a548d7c35b5JmltdHM9MTc3MjQ5NjAwMA&ptn=3&ver=2&hsh=4&fclid=10287227-ff9f-64a2-05b4-6535fe2065b2&u=a1aHR0cHM6Ly9uZXdzLm1pdC5lZHUvMjAyNS9tYWtpbmctY2xlYW4tZW5lcmd5LWludmVzdG1lbnRzLW1vcmUtc3VjY2Vzc2Z1bC0xMjEy&ntb=1

Making clean energy investments more successful - MIT News

Dec 12, 2025 · New research emphasizes the importance of well-validated models and forecasting tools in evaluating choices for investments in clean energy technologies and policies by governments and …

arxiv.org/abs/2205.05805v1

SubER: A Metric for Automatic Evaluation of Subtitle Quality

This paper addresses the problem of evaluating the quality of automatically generated subtitles, which includes not only the quality of the machine-transcribed or translated speech, but also the quality of line segmentation and subtitle timing. We pr...

arxiv.org/abs/1909.11995v3

The Stroke Correspondence Problem, Revisited

We revisit the stroke correspondence problem [13,14]. We optimize this algorithm by 1) evaluating suitable preprocessing (normalization) methods 2) extending the algorithm with an additional distance measure to handle Hiragana, Katakana and Kanji cha...

arxiv.org/abs/hep-ex/0305078v1

Evaluating current processors performance and machines stability

Accurately estimate performance of currently available processors is becoming a key activity, particularly in HENP environment, where high computing power is crucial. This document describes the methods and programs, opensource or freeware, used to...

arxiv.org/abs/astro-ph/9811273v1

Igloo Pixelizations

Upcoming microwave background experiments will see an incredible increase in the volume of data to be analyzed, which makes the choice of how it is discretized on the sky a crucial issue. I discuss criteria for evaluating different pixelizations an...

arxiv.org/abs/2308.02490v4

MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities

We propose MM-Vet, an evaluation benchmark that examines large multimodal models (LMMs) on complicated multimodal tasks. Recent LMMs have shown various intriguing abilities, such as solving math problems written on the blackboard, reasoning about eve...