arxiv.org/abs/1809.07567v1
Mobile phone data are an interesting new data source for official statistics. However, multiple problems and uncertainties need to be solved before these data can inform, support or even become an integral part of statistical production processes. In...
www.bing.com/ck/a?!&&p=651536dc44bde2f8ecf0b3f972c7ce0f74025abe71c83d3f0d992d65b4690ab1JmltdHM9MTc3MjE1MDQwMA&ptn=3&ver=2&hsh=4&fclid=2554f076-2da4-6ea2-12c1-e77b2cb36f6b&u=a1aHR0cHM6Ly93d3cuYmVsbW9udGZvcnVtLm9yZy93cC1jb250ZW50L3VwbG9hZHMvMjAxOS8xMC9DUkFfRGF0YV9EaWdpdGFsX091dHB1dHNfTWFuYWdlbWVudF9WMi5wZGY&ntb=1
A full Data and Digital Outputs Management Plan for an awarded Belmont Forum project is a living, actively updated document that describes the data management life cycle for the data and other …
www.bing.com/ck/a?!&&p=ded18e70e31d9bee1b596cedf6ac2b7f23b3a4263170189126f6447c76f02415JmltdHM9MTc3MjE1MDQwMA&ptn=3&ver=2&hsh=4&fclid=2554f076-2da4-6ea2-12c1-e77b2cb36f6b&u=a1aHR0cHM6Ly93d3cuYmVsbW9udGZvcnVtLm9yZy93cC1jb250ZW50L3VwbG9hZHMvMjAxOS8xMC9DdXJyLURldi0yMDE3MDQyOF9BMi1CaXNob3AucGRm&ntb=1
Apr 28, 2017 · Several actions related to the data lifecycle, such as data discovery, do require an understanding of the data, technology, and information infrastructures that may result from …
arxiv.org/abs/1603.04395v1
As competitions get more popular, transferring ever-larger data sets becomes infeasible and costly. For example, downloading the 157.3 GB 2012 ImageNet data set incurs about $4.33 in bandwidth costs per download. Downloading the full ImageNet data se...
arxiv.org/abs/2504.08364v2
Selecting data points for model training is critical in machine learning. Effective selection methods can reduce the labeling effort, optimize on-device training for embedded systems with limited data storage, and enhance the model performance. This...
arxiv.org/abs/1501.07329v4
Model-based analysis tools, built on assumptions and simplifications, are difficult to handle smart grids with data characterized by 4Vs data. This paper, using random matrix theory (RMT), motivates data-driven tools to perceive the complex grids in...
arxiv.org/abs/2101.00706v3
Autonomous vehicles require fleet-wide data collection for continuous algorithm development and validation. The Smart Black Box (SBB) intelligent event data recorder has been proposed as a system for prioritized high-bandwidth data capture. This pape...
arxiv.org/abs/2306.04338v1
Data science has become increasingly essential for the production of official statistics, as it enables the automated collection, processing, and analysis of large amounts of data. With such data science practices in place, it enables more timely, mo...
www.bing.com/ck/a?!&&p=ba63bf52c1617ce17b0e245c8e5390f23b179162024e45cff893b9b664bf271dJmltdHM9MTc3MjE1MDQwMA&ptn=3&ver=2&hsh=4&fclid=0ceef642-f838-6d2a-1bd6-e14ff9ab6c2b&u=a1aHR0cHM6Ly9zdXBwb3J0Lmdvb2dsZS5jb20vZG9jcy9hbnN3ZXIvMzA5MzM0Mz9obD16aC1IYW50&ntb=1
In case of mixed data types in a single column, the majority data type determines the data type of the column for query purposes. Minority data types are considered null values. query - 要執行的查詢作業 …
www.bing.com/ck/a?!&&p=93be114f0f97427e400b8291ba24ce50279d9ae514301d3fb52a17f51adb1103JmltdHM9MTc3MjE1MDQwMA&ptn=3&ver=2&hsh=4&fclid=3eea4e4b-e755-6cdb-2ebf-5946e6156dd4&u=a1aHR0cHM6Ly9zdXBwb3J0Lmdvb2dsZS5jb20vZG9jcy9hbnN3ZXIvMzA5MzM0Mz9obD16aC1IYW50&ntb=1
In case of mixed data types in a single column, the majority data type determines the data type of the column for query purposes. Minority data types are considered null values. query - 要執行的查詢作業 …
arxiv.org/abs/2508.00932v1
Previous literature has proposed that the companies operating data centers enforce government regulations on AI companies. Using a new dataset of 775 non-U.S. data center projects, this paper estimates how often data centers could be subject to forei...
github.com/paulcheng0830/Data-Analysis-Projects---Hong-Kong-Real-Estate-Market-Data-Analysis-2020
No description (⭐ 0)
arxiv.org/abs/2110.02311v2
While India has been one of the hotspots of COVID-19, data about the pandemic from the country has proved to be largely inaccessible at scale. Much of the data exists in unstructured form on the web, and limited aspects of such data are available thr...
arxiv.org/abs/1609.02137v1
Various methods to automate traffic data collection have recently been developed by many researchers. A macroscopic data collection through image processing has been proposed. For microscopic traffic flow data, such as individual speed and time or di...
arxiv.org/abs/1004.4718v1
In this paper, we emphasize the need for data cleansing when clustering large-scale transaction databases and propose a new data cleansing method that improves clustering quality and performance. We evaluate our data cleansing method through a series...
arxiv.org/abs/1805.00519v1
Social media users generate tremendous amounts of data. To better serve users, it is required to share the user-related data among researchers, advertisers and application developers. Publishing such data would raise more concerns on user privacy. To...
www.bing.com/ck/a?!&&p=c922cfbd30fa9abd2d49e9f1552a2cec4d10cd61bdce7c994cb734dab80d6f17JmltdHM9MTc3MjA2NDAwMA&ptn=3&ver=2&hsh=4&fclid=28ae9124-6080-65f1-3615-8628611364a3&u=a1aHR0cHM6Ly93d3cuemFja3MuY29tL3N0b2NrL3F1b3RlL2FkcA&ntb=1
1 day ago · View Automatic Data Processing, Inc ADP investment & stock information. Get the latest Automatic Data Processing, Inc ADP detailed stock quotes, stock data, Real-Time ECN, charts, …
arxiv.org/abs/2511.11755v1
The fragmentation of public data in Brazil, coupled with inconsistent standards and limited interoperability, hinders effective research, evidence-based policymaking and access to data-driven insights. To address these issues, we introduce Brazil Dat...
arxiv.org/abs/1603.06828v2
Revealing hidden geometry and topology in noisy data sets is a challenging task. Elastic principal graph is a computationally efficient and flexible data approximator based on embedding a graph into the data space and minimizing the energy functional...
arxiv.org/abs/1204.3055v1
We present a data model describing the structure of spectrophotometric datasets with spectral and temporal coordinates and associated metadata. This data model may be used to represent spectra, time series data, segments of SED (Spectral Energy Distr...