arxiv.org/abs/2311.03386v1
Data attribution methods play a crucial role in understanding machine learning models, providing insight into which training data points are most responsible for model outputs during deployment. However, current state-of-the-art approaches require a...
arxiv.org/abs/2011.11355v3
In this paper, we present a data-driven controller design method for continuous-time nonlinear systems, using no model knowledge but only measured data affected by noise. While most existing approaches focus on systems with polynomial dynamics, our a...
github.com/tylerjrichards/Streamlit-for-Data-Science
A repo for the book 'Streamlit for Data Science' by Tyler Richards (⭐ 245)
arxiv.org/abs/2407.08113v1
Dataset distillation synthesizes a small set of images from a large-scale real dataset such that synthetic and real images share similar behavioral properties (e.g, distributions of gradients or features) during a training process. Through extensive...
github.com/podaac/data-subscriber
Subscribe and bulk download collections of data at PO.DAAC (⭐ 114)
arxiv.org/abs/2412.09152v1
Recent research is revealing data-sonification as a promising complementary approach to vision, benefiting both data perception and interpretation. We present herakoi, a novel open-source software that uses machine learning to allow real-time image s...
www.reddit.com/r/DataHoarder/comments/1ray20i/so_what_happens_if_archivetoday_goes_go_down/
Context for the unaware: [https://www.reddit.com/r/DataHoarder/comments/1qspk6x/archivetoday\_is\_directing\_a\_ddos\_attack\_against/](https://www.reddit.com/r/DataHoarder/comments/1qspk6x/archivetod...
arxiv.org/abs/1708.00195v1
Context: The first Gaia data release (DR1) delivered a catalogue of astrometry and photometry for over a billion astronomical sources. Within the panoply of methods used for data exploration, visualisation is often the starting point and even the gui...
en.wikipedia.org/wiki/Data_analysis
that analyzes data about customer purchase history, and uses the results to recommend other purchases the customer might enjoy. Once data is analyzed, it
arxiv.org/abs/1802.05219v1
This paper presents a framework for generating adventure games from open data. Focusing on the murder mystery type of adventure games, the generator is able to transform open data from Wikipedia articles, OpenStreetMap and images from Wikimedia Commo...
github.com/parth-psd-009/Instagram-Lok-Sabha-Data-Analysis
This is a work on the data analysis of the Instagram pages of various Lok Sabha members in the 2024 elections. (⭐ 0)
arxiv.org/abs/2110.01056v1
Collaboration across institutional boundaries is widespread and increasing today. It depends on federations sharing data that often have governance rules or external regulations restricting their use. However, the handling of data governance rules (a...
arxiv.org/abs/1008.1188v1
The basic objective of data visualization is to provide an efficient graphical display for summarizing and reasoning about quantitative information. During the last decades, political science has accumulated a large corpus of various kinds of data su...
arxiv.org/abs/2211.14769v4
Federated embodied agent learning protects the data privacy of individual visual environments by keeping data locally at each client (the individual environment) during training. However, since the local data is inaccessible to the server under feder...
github.com/BloombergGraphics/2024-h1b-immigration-data
US H-1B Visa Lottery and Petition Data FY 2021 - FY 2024 (⭐ 233)
github.com/Apress/pro-hadoop-data-analytics
Source code for 'Pro Hadoop Data Analytics' by Kerry Koitzsch (⭐ 14)
en.wikipedia.org/wiki/Microsoft_Azure_SQL_Database
Windows Azure SQL Database) is a managed cloud database (PaaS) cloud-based Microsoft SQL Servers, provided as part of Microsoft Azure services. The service
github.com/NVlabs/ffhq-dataset
Flickr-Faces-HQ Dataset (FFHQ) (⭐ 4107)
github.com/pandas-dev/pandas
Flexible and powerful data analysis / manipulation library for Python, providing labeled data structures similar to R data.frame objects, statistical functions, and much more (⭐ 48073)
arxiv.org/abs/1510.06871v8
We present the R-package mgm for the estimation of k-order Mixed Graphical Models (MGMs) and mixed Vector Autoregressive (mVAR) models in high-dimensional data. These are a useful extensions of graphical models for only one variable type, since data...