arxiv.org/abs/2004.12341v1
In this article, we develop a data assimilation procedure to predict the evolution of epidemics with data uncertainty, with application to the Covid-19 pandemic. We construct a vademecum of solutions by solving the SIR epidemic model for a set of dat...
arxiv.org/abs/2206.14414v1
In recent years, we have witnessed an explosive growth of data. Much of this data is video data generated by security cameras, smartphones, and dash cams. The timely analysis of such data is of great practical importance for many emerging application...
arxiv.org/abs/2108.00319v4
Artifacts in functional MRI (fMRI) data cause deviations from common distributional assumptions, introduce spatial and temporal outliers, and reduce the signal-to-noise ratio of the data -- all of which can have negative consequences for downstream s...
www.bing.com/ck/a?!&&p=ff190df3e8d964edcb1ed306f2389c00f5c067558d34ec01f3b9ede953cc49b4JmltdHM9MTc3MjQ5NjAwMA&ptn=3&ver=2&hsh=4&fclid=29e5a191-a6ac-647c-2428-b683a746651d&u=a1aHR0cHM6Ly9zdXBwb3J0Lmdvb2dsZS5jb20vZG9jcy9hbnN3ZXIvMzA5MzM0Mz9obD16aC1IYW50&ntb=1
In case of mixed data types in a single column, the majority data type determines the data type of the column for query purposes. Minority data types are considered null values. query - 要執行的查詢作業 …
arxiv.org/abs/1505.04935v2
Continued reliance on human operators for managing data centers is a major impediment for them from ever reaching extreme dimensions. Large computer systems in general, and data centers in particular, will ultimately be managed using predictive compu...
github.com/timothycarambat/senate-stock-watcher-data
Data repository of JSON files that are filed by US Senators on efdsearch.senate.gov where they must report their stock trades. This is the same data as on senatestockwatcher.com (⭐ 69)
arxiv.org/abs/0907.2471v1
Declarative data quality has been an active research topic. The fundamental principle behind a declarative approach to data quality is the use of declarative statements to realize data quality primitives on top of any relational data source. A prim...
arxiv.org/abs/1902.01304v1
The area of declarative data analytics explores the application of the declarative paradigm on data science and machine learning. It proposes declarative languages for expressing data analysis tasks and develops systems which optimize programs writte...
arxiv.org/abs/1011.3344v2
The onboard software and data communication in the RT-2 Experiment onboard the Coronas-Photon satellite is organized in a hierarchical way to effectively handle and communicate asynchronous data generated by the X-ray detectors. A flexible data handl...
arxiv.org/abs/2502.14854v2
LLM developers are increasingly reliant on synthetic data, but generating high-quality data for complex long-context reasoning tasks remains challenging. We introduce CLIPPER, a compression-based approach for generating synthetic data tailored to nar...
arxiv.org/abs/2409.03741v1
Machine learning has revolutionized numerous domains, playing a crucial role in driving advancements and enabling data-centric processes. The significance of data in training models and shaping their performance cannot be overstated. Recent research...
arxiv.org/abs/1611.00481v2
In the era of big data, it is common to have data with multiple modalities or coming from multiple sources, known as "multi-view data". Multi-view clustering provides a natural way to generate clusters from such data. Since different views share some...
www.bing.com/ck/a?!&&p=3ba2a2c93fe88aba6e702908d78c01b178014584be60d2c6f4110ffc74479e53JmltdHM9MTc3MjQ5NjAwMA&ptn=3&ver=2&hsh=4&fclid=2e84aa39-ee0f-6170-0fcc-bd2bef3a605c&u=a1aHR0cHM6Ly9lbmdsaXNoLnN0YWNrZXhjaGFuZ2UuY29tL3F1ZXN0aW9ucy8xNjEwODEvd2hhdC1pcy1mcmVlLWZvcm0tZGF0YS1lbnRyeQ&ntb=1
If you are storing documents, however, you should choose either the mediumtext or longtext type. Could you please tell me what free-form data entry is? I know what data entry is per se - when data is fed …
arxiv.org/abs/1601.03115v1
Big Data can mean different things to different people. The scale and challenges of Big Data are often described using three attributes, namely Volume, Velocity and Variety (3Vs), which only reflect some of the aspects of data. In this chapter we rev...
arxiv.org/abs/2105.03161v1
In the Open Data Portal Germany (OPAL) project, a pipeline of the following data refinement steps has been developed: requirements analysis, data acquisition, analysis, conversion, integration and selection. 800,000 datasets in DCAT format have been...
arxiv.org/abs/1811.01429v2
Multivariate functional data are becoming ubiquitous with advances in modern technology and are substantially more complex than univariate functional data. We propose and study a novel model for multivariate functional data where the component proces...
arxiv.org/abs/2509.10165v2
Companies are looking to data anonymization research $\unicode{x2013}$ including differential private and synthetic data methods $\unicode{x2013}$ for simple and straightforward compliance solutions. But data anonymization has not taken off in practi...
arxiv.org/abs/1302.4133v1
NVD is one of the most popular databases used by researchers to conduct empirical research on data sets of vulnerabilities. Our recent analysis on Chrome vulnerability data reported by NVD has revealed an abnormally phenomenon in the data where almos...
github.com/ift-gftc/SeafoodTrackathon
A hackathon using real data sets taken from industry supply chains to develop solutions which address tools and applications to capture data on fishing vessels, conceptualize feasible data sharing options and develop creative ways to define data formats that s…
arxiv.org/abs/2308.16109v1
Accessibility of research data is critical for advances in many research fields, but textual data often cannot be shared due to the personal and sensitive information which it contains, e.g names or political opinions. General Data Protection Regulat...