r - Rounding selected columns of data.table - Stack Overflow
Jan 26, 2017 · I have following data and code to round selected columns of this data.table: mydf = structure (list (vnum1 = c (0.590165705411504, -1.39939534199836, 0.720226053660755, …
Jan 26, 2017 · I have following data and code to round selected columns of this data.table: mydf = structure (list (vnum1 = c (0.590165705411504, -1.39939534199836, 0.720226053660755, …
Fog Reveal is a tracking tool that aggregates location data from mobile apps. It is a product of FOG Data Science. FOG Data Science is a limited liability
Identifying correspondences in noisy data is a critically important step in estimation processes. When an informative initial estimation guess is available, the data association challenge is less acute; however, the existence of a high-quality initia...
Current pre-trained language models (PLM) are typically trained with static data, ignoring that in real-world scenarios, streaming data of various sources may continuously grow. This requires PLMs to integrate the information from all the sources in...
Recreation of Cole Nussbaumer Knaflic's Storytelling with Data plots using R an ggplot2 (⭐ 233)
The lack of data regarding Information and Communications Technology sector alumni data is a known problem in several countries including Egypt. It is not clear what entry and senior jobs are occupied by alumni and which countries attract them. This...
The easiest way to waste your data. (⭐ 87)
In this paper we present a workflow management system which permits the kinds of data-driven workflows required by urgent computing, namely where new data is integrated into the workflow as a disaster progresses in order refine the predictions as tim...
Emerging connected vehicle (CV) data sets have recently become commercially available. This paper presents several tools using CV data to evaluate traffic progression quality along a signalized corridor. These include both performance measures for hi...
Data represented as strings abounds in biology, linguistics, document mining, web search and many other fields. Such data often have a hierarchical structure, either because they were artificially designed and composed in a hierarchical manner or bec...
Often logs hosted in large data centers represent network traffic data over a long period of time. For instance, such network traffic data logged via a TCP dump packet sniffer (as considered in the 1998 DARPA intrusion attack) included network packet...
In computer science, a succinct data structure is a data structure which uses an amount of space that is "close" to the information-theoretic lower bound
In computer science, a purely functional data structure is a data structure that can be directly implemented in a purely functional language. The main
Traditional data science education often omits training on research workflows: the process that moves a scientific investigation from raw data to coherent research question to insightful contribution. In this paper, we elaborate basic principles of a...
Today's big data science communities manage their data publication and replication at the application layer. These communities utilize myriad mechanisms to publish, discover, and retrieve datasets - the result is an ecosystem of either centralized, o...
This paper introduces Ali-AUG, a novel single-step diffusion model for efficient labeled data augmentation in industrial applications. Our method addresses the challenge of limited labeled data by generating synthetic, labeled images with precise fea...
Welcome to Optery’s open-source directory of data brokers and opt-out information, the largest of its kind. (⭐ 45)
Data dumps from thesession.org (⭐ 81)
Free football data from StatsBomb (⭐ 3029)
The count-min sketch (CMS) is a time and memory efficient randomized data structure that provides estimates of tokens' frequencies in a data stream of tokens, i.e. point queries, based on random hashed data. A learning-augmented version of the CMS, r...