860 results for Evaluating · 0.094s

arxiv.org/abs/2207.12467v1

Reproducible Sorbent Materials Foundry for Carbon Capture at Scale

We envision an autonomous sorbent materials foundry (SMF) for rapidly evaluating materials for direct air capture of carbon dioxide (CO2), specifically targeting novel metal organic framework materials. Our proposed SMF is hierarchical, simultaneousl...

Sponsored Partners
www.bing.com/ck/a?!&&p=9bc6876254301ae668511d575a27ebab22041726eb7a8b09d29d38563388a47eJmltdHM9MTc3MjQ5NjAwMA&ptn=3&ver=2&hsh=4&fclid=0c9a1514-eb06-655e-29da-0206ea59642c&u=a1aHR0cHM6Ly93d3cuYWNyb2Jpb3N5c3RlbXMuY29tL2luc2lnaHRzLzI0Mjk&ntb=1

Human Erythropoietin and its use in LNP-mediated mRNA Drug …

Mar 15, 2024 · When evaluating LNP-mediated mRNA delivery efficacy and putatively pharmacologic effects of drugs in vivo / in vitro, the use of human EPO (hEPO) -encoding …

arxiv.org/abs/2504.02463v1

Evaluating AI Recruitment Sourcing Tools by Human Preference

This study introduces a benchmarking methodology designed to evaluate the performance of AI-driven recruitment sourcing tools. We created and utilized a dataset to perform a comparative analysis of search results generated by leading AI-based solutio...

www.bing.com/ck/a?!&&p=4689b1ae860b0624b3f3174c883ec465034b7edd895247233ee22970f5bf8c87JmltdHM9MTc3MjQ5NjAwMA&ptn=3&ver=2&hsh=4&fclid=36107c4d-e80d-6360-0d42-6b5fe93862d4&u=a1aHR0cHM6Ly93d3cuZGVzbW9zLmNvbS9zY2llbnRpZmlj&ntb=1

Scientific Calculator - Desmos

A beautiful, free online scientific calculator with advanced features for evaluating percentages, fractions, exponential functions, logarithms, trigonometry, statistics, and more.

arxiv.org/abs/1912.06101v1

The PlayStation Reinforcement Learning Environment (PSXLE)

We propose a new benchmark environment for evaluating Reinforcement Learning (RL) algorithms: the PlayStation Learning Environment (PSXLE), a PlayStation emulator modified to expose a simple control API that enables rich game-state representations. W...

arxiv.org/abs/2406.04662v1

Evaluating and Mitigating IP Infringement in Visual Generative AI

The popularity of visual generative AI models like DALL-E 3, Stable Diffusion XL, Stable Video Diffusion, and Sora has been increasing. Through extensive evaluation, we discovered that the state-of-the-art visual generative models can generate conten...

github.com/facebookresearch/SentEval

facebookresearch/SentEval

A python tool for evaluating the quality of sentence embeddings. (⭐ 2106)

arxiv.org/abs/1301.7383v1

Evaluating Las Vegas Algorithms - Pitfalls and Remedies

Stochastic search algorithms are among the most sucessful approaches for solving hard combinatorial problems. A large class of stochastic search approaches can be cast into the framework of Las Vegas Algorithms (LVAs). As the run-time behavior of LVA...

arxiv.org/abs/2410.12784v2

JudgeBench: A Benchmark for Evaluating LLM-based Judges

LLM-based judges have emerged as a scalable alternative to human evaluation and are increasingly used to assess, compare, and improve models. However, the reliability of LLM-based judges themselves is rarely scrutinized. As LLMs become more advanced,...

arxiv.org/abs/2507.20985v1

Behavioral Study of Dashboard Mechanisms

Visualization dashboards are increasingly used in strategic settings like auctions to enhance decision-making and reduce strategic confusion. This paper presents behavioral experiments evaluating how different dashboard designs affect bid optimizatio...

github.com/shayneobrien/explicit-gan-eval

shayneobrien/explicit-gan-eval

Code for reproducing the results of "Evaluating Generative Adversarial Networks on Explicitly Parameterized Distributions" (O'Brien et al., NeurIPS 2018). (⭐ 10)