arxiv.org/abs/2601.19176v1
Data lakes have emerged as a flexible and scalable solution for storing and analyzing large volumes of heterogeneous data, including structured, semi-structured, and unstructured formats. Despite their growing adoption in both industry and academia,...
github.com/aws-solutions-library-samples/data-lakes-on-aws
Enterprise-grade, production-hardened, serverless data lake on AWS (⭐ 479)
github.com/Hack-Education-Data/emerson-collective
This repository contains data about the venture philanthropy firm Emerson Collective (⭐ 2)
www.bing.com/ck/a?!&&p=317a83e71910bf9be10b45ff670dc523e33e9658e12a0199fa71e56b999561d0JmltdHM9MTc3MjY2ODgwMA&ptn=3&ver=2&hsh=4&fclid=24a7f6aa-f8b8-6e0d-27c1-e1b9f9716ff9&u=a1aHR0cHM6Ly95b3Vnb3YuY29tL2VuLXVzL2NvbnRlbnQ&ntb=1
Access free public data, surveys, articles, trackers and rankings. Explore YouGov’s comprehensive research that combines real-world data with expert analysis.
arxiv.org/abs/2602.18588v1
Managing the data and metadata during the active development phase of an experimental project presents a significant challenge, particularly in collaborative research. This phase is frequently overlooked in Data Management Plans included in project p...
arxiv.org/abs/1506.03136v1
The General Single-Dish Data format (GSDD) was developed in the mid-1980s as a data model to support centimeter, millimeter and submillimeter instrumentation at NRAO, JCMT, the University of Arizona and IRAM. We provide an overview of the GSDD requir...
www.bing.com/ck/a?!&&p=ba3a9371b9f852ae8480906ad02b519142c61d577cbe5cf9470b0aba383a559fJmltdHM9MTc3MjY2ODgwMA&ptn=3&ver=2&hsh=4&fclid=27cb83e4-e2bf-689a-00ea-94f7e37069a5&u=a1aHR0cHM6Ly9zdXBlcnVzZXIuY29tL3F1ZXN0aW9ucy80MTk4MzIvaG93LWNhbi1pLW9wZW4tdGhlLTMyLWJpdC1vZGJjLWRhdGEtc291cmNlLWFkbWluaXN0cmF0b3ItaW4td2luZG93cy03LTY0LWJpdA&ntb=1
I want to add 32-bit data sources. There seems to be no obvious way to see existing instances of these or create new ones. How can I open the 32-bit "ODBC Data Source Administrator" window in Wind...
arxiv.org/abs/1808.01883v3
News about massive online breaches is increasingly common. But there has been little good data on how exposed people are because of these breaches. We combine data from a large, representative sample of adult Americans (n = 5,000) with data from \tex...
github.com/LinkedInLearning/advanced-python-working-with-data-4312001
This is a repository for the LinkedIn Learning course Advanced Python: Working With Data (⭐ 100)
arxiv.org/abs/2401.08895v4
The input data pipeline is an essential component of each machine learning (ML) training job. It is responsible for reading massive amounts of training data, processing batches of samples using complex transformations, and loading them onto training...
arxiv.org/abs/2303.02761v1
One of the factors limiting the performance of handwritten text recognition (HTR) for stenography is the small amount of annotated training data. To alleviate the problem of data scarcity, modern HTR methods often employ data augmentation. However, d...
github.com/talos/nypd-crash-data-bandaid
nypd traffic crash data has a booboo. this eases the pain. (⭐ 41)
github.com/Apress/high-impact-data-visualization-in-excel
Source code for 'High Impact Data Visualization in Excel with Power View, 3D Maps, Get & Transform and Power BI' by Adam Aspin (⭐ 8)
github.com/kiseki1107/College-Scorecard-Data-Analysis
Provided publicly by the U.S. Department of Education's College Scorecard, this data visualization project aims to identify factors that may lead to student success post-undergraduate completion. (⭐ 10)
github.com/AnishaShelke/Periodical-Magazines-Digital-Editions
Whether to inform something or entertain someone, media is the first to hit your mind. Until last decade, printing of a magazine, booklet or a newsletter was done and a hard copy was available. Today, in this internet era, it has become much easier to share th…
arxiv.org/abs/2304.02189v1
The interactive exploration of large and evolving datasets is challenging as relationships between underlying variables may not be fully understood. There may be hidden trends and patterns in the data that are worthy of further exploration and analys...
stackoverflow.com/questions/14262433/large-data-workflows-using-pandas
Tags: python, mongodb, pandas, hdf5, large-data | Score: 1190
stackoverflow.com/questions/16013207/storing-non-valuable-data-in-ios
Tags: ios, objective-c, core-data | Score: 0
www.bing.com/ck/a?!&&p=d2c37fb4080ced69e53dd7f1ea1d8ef2693b479be96698673efd92dbb3e4b322JmltdHM9MTc3MjY2ODgwMA&ptn=3&ver=2&hsh=4&fclid=16417ef1-d839-60cb-2b56-69e2d9ae61d2&u=a1aHR0cHM6Ly93d3cuY2RjLmdvdi9wbGFjZXMvaW5kZXguaHRtbA&ntb=1
Access current PLACES data by county, place, census tract, and ZIP Code tabulation area. Access and download local data from 2016-2019 CDC's 500 Cities project (precursor to PLACES). Explore the …
stackoverflow.com/questions/48970967/is-data-attributes-posing-any-security-issues
Tags: html, security, custom-data-attribute | Score: 1