860 results for Evaluating · 0.097s

arxiv.org/abs/1302.5549v1

On Graph Deltas for Historical Queries

In this paper, we address the problem of evaluating historical queries on graphs. To this end, we investigate the use of graph deltas, i.e., a log of time-annotated graph operations. Our storage model maintains the current graph snapshot and the delt...

www.bing.com/ck/a?!&&p=d6c57a447aee2726d9daabcd3bd8015431aaf47968d9c4be8ee1637f5b7b3830JmltdHM9MTc3MjU4MjQwMA&ptn=3&ver=2&hsh=4&fclid=2d8004d0-b54e-6330-157a-13c2b40a6288&u=a1aHR0cHM6Ly9tZWFzdXJlLXdlbGxiZWluZy5vcmcvd2VsbGJlaW5nLWV4cGxhaW5lZC8&ntb=1

What is wellbeing, and what matters? – Evaluating wellbeing

One way of understanding wellbeing is how well people are able to flourish – whether they feel positive emotions, can function well in society, can respond to challenges and make meaning in their lives – …

Sponsored Partners
arxiv.org/abs/1911.07707v1

Building Fast Fuzzers

Fuzzing is one of the key techniques for evaluating the robustness of programs against attacks. Fuzzing has to be effective in producing inputs that cover functionality and find vulnerabilities. But it also has to be efficient in producing such input...

arxiv.org/abs/1704.04579v1

Evaluating Quality of Chatbots and Intelligent Conversational Agents

Chatbots are one class of intelligent, conversational software agents activated by natural language input (which can be in the form of text, voice, or both). They provide conversational output in response, and if commanded, can sometimes also execute...

arxiv.org/abs/1011.4778v2

Evaluating cumulative ascent: Mountain biking meets Mandelbrot

The problem of determining total distance ascended during a mountain bike trip is addressed. Altitude measurements are obtained from GPS receivers utilizing both GPS-based and barometric altitude data, with data averaging used to reduce fluctuations....

arxiv.org/abs/2412.18204v2

BoxMAC -- A Boxing Dataset for Multi-label Action Classification

In competitive combat sports like boxing, analyzing a boxers's performance statics is crucial for evaluating the quantity and variety of punches delivered during bouts. These statistics provide valuable data and feedback, which are routinely used for...

arxiv.org/abs/2506.15740v2

SHADE-Arena: Evaluating Sabotage and Monitoring in LLM Agents

As Large Language Models (LLMs) are increasingly deployed as autonomous agents in complex and long horizon settings, it is critical to evaluate their ability to sabotage users by pursuing hidden objectives. We study the ability of frontier LLMs to ev...

arxiv.org/abs/2401.02984v3

Large Language Models in Mental Health Care: a Scoping Review

Objectieve:This review aims to deliver a comprehensive analysis of Large Language Models (LLMs) utilization in mental health care, evaluating their effectiveness, identifying challenges, and exploring their potential for future application. Materials...

www.bing.com/ck/a?!&&p=f085c9a43a69cf6fc3a3cf540095a7f80a4de80329c5209f7087e6c4608bd008JmltdHM9MTc3MjU4MjQwMA&ptn=3&ver=2&hsh=4&fclid=113e27be-9842-6a18-32ce-30ac992a6b2e&u=a1aHR0cHM6Ly93d3cuZGVzbW9zLmNvbS9zY2llbnRpZmlj&ntb=1

Scientific Calculator - Desmos

A beautiful, free online scientific calculator with advanced features for evaluating percentages, fractions, exponential functions, logarithms, trigonometry, statistics, and more.

arxiv.org/abs/2205.14705v1

Evaluating the Socioeconomic Status of a Large Social Event Attendees

In this study, Call Detail Records (CDRs) from downtown Budapest were analysed, focusing on a large-scale event in August 2014. The attendees of the main event of the Hungarian State Foundation Day have been analysed based on their Socioeconomic Stat...

github.com/aesara-devs/aesara

aesara-devs/aesara

Aesara is a Python library for defining, optimizing, and efficiently evaluating mathematical expressions involving multi-dimensional arrays. (⭐ 1218)

arxiv.org/abs/2409.10566v1

Eureka: Evaluating and Understanding Large Foundation Models

Rigorous and reproducible evaluation is critical for assessing the state of the art and for guiding scientific advances in Artificial Intelligence. Evaluation is challenging in practice due to several reasons, including benchmark saturation, lack of...

www.reddit.com/r/bigseo/comments/4m32gf/tools_for_competitive_keyword_research/

Tools for Competitive Keyword Research

Hey fellow SEO'rs, I'm hoping to get some feedback on any tools or platforms you uses for competitive keyword research. We used to use MOZ's platform, but we're now in the process of evaluating wheth...