1,230 results for Datasets · 0.106s

Sponsored Partners
arxiv.org/abs/1809.08980v1

A Systematic Study of Hale and Anti-Hale Sunspot Physical Parameters

We present a systematic study of sunspot physical parameters using full disk magnetograms from MDI/SoHO and HMI/SDO. Our aim is to use uniform datasets and analysis procedures to characterize the sunspots, paying particular attention to the differenc...

arxiv.org/abs/2303.01949v1

Artificial Intelligence for Dementia Research Methods Optimization

Introduction: Machine learning (ML) has been extremely successful in identifying key features from high-dimensional datasets and executing complicated tasks with human expert levels of accuracy or greater. Methods: We summarize and critically evaluat...

github.com/shreyans29/thesemicolon

shreyans29/thesemicolon

This repository contains Ipython notebooks and datasets for the data analytics youtube tutorials on The Semicolon. (⭐ 395)

github.com/google-research/meta-dataset

google-research/meta-dataset

A dataset of datasets for learning to learn from few examples (⭐ 799)

github.com/Amaan-29/Dow_Jones_Industrial_Average_Prediction

Amaan-29/Dow_Jones_Industrial_Average_Prediction

GROUP PROJECT Context: Dow Jones Industrial Average (DJIA) prediction We will be predicting the DJIA closing value by using the top 25 headlines for the day. The Dow Jones Industrial Average, Dow Jones, or simply the Dow, is a stock market index that measures the stock performanc…

arxiv.org/abs/2205.15661v1

NEWTS: A Corpus for News Topic-Focused Summarization

Text summarization models are approaching human levels of fidelity. Existing benchmarking corpora provide concordant pairs of full and abridged versions of Web, news or, professional content. To date, all summarization datasets operate under a one-si...

arxiv.org/abs/1901.00850v2

CLEVR-Ref+: Diagnosing Visual Reasoning with Referring Expressions

Referring object detection and referring image segmentation are important tasks that require joint understanding of visual information and natural language. Yet there has been evidence that current benchmark datasets suffer from bias, and current sta...

arxiv.org/abs/1709.08001v1

Real-time Log Query Interface for large datasets using Apache Spark

Log Query Interface is an interactive web application that allows users to query the very large data logs of MobileInsight easily and efficiently. With this interface, users no longer need to talk to the database through command line queries, nor to...

github.com/hsnr-data-science/SEDAR

hsnr-data-science/SEDAR

A Semantic Data Reservoir for Heterogeneous Datasets (⭐ 8)

arxiv.org/abs/1211.0496v1

W/Z + jet Production at the LHC

This paper summarises results on W and Z plus jet production in pp collisions at $\sqrt{s} = 7$ TeV at the CERN Large Hadron Collider, from both the ATLAS and CMS experiments. Based on the 2010 and 2011 datasets, measurements have been made of numero...

github.com/bids-standard/pybids

bids-standard/pybids

Python tools for querying and manipulating BIDS datasets. (⭐ 256)

github.com/bids-standard/bids-examples

bids-standard/bids-examples

A set of BIDS compatible datasets with empty raw data files that can be used for writing lightweight software tests. (⭐ 213)

github.com/candlewill/Dialog_Corpus

candlewill/Dialog_Corpus

用于训练中英文对话系统的语料库 Datasets for Training Chatbot System (⭐ 2051)

arxiv.org/abs/2407.03550v2

CoMix: A Comprehensive Benchmark for Multi-Task Comic Understanding

The comic domain is rapidly advancing with the development of single-page analysis and synthesis models. However, evaluation metrics and datasets lag behind, often limited to small-scale or single-style test sets. We introduce a novel benchmark, CoMi...

arxiv.org/abs/1809.00027v1

Dimensionality-Reduction of Climate Data using Deep Autoencoders

We explore the use of deep neural networks for nonlinear dimensionality reduction in climate applications. We train convolutional autoencoders (CAEs) to encode two temperature field datasets from pre-industrial control runs in the CMIP5 first ensembl...

arxiv.org/abs/2305.13026v2

DUMB: A Benchmark for Smart Evaluation of Dutch Models

We introduce the Dutch Model Benchmark: DUMB. The benchmark includes a diverse set of datasets for low-, medium- and high-resource tasks. The total set of nine tasks includes four tasks that were previously not available in Dutch. Instead of relying...

github.com/oleiade/Elevator

oleiade/Elevator

Elevator is an open source, on-disk key-value store. Provides high-performance bulk read-write operations over very large datasets while exposing a simple and efficient API. (⭐ 70)

github.com/RiccardoRiccio/Fitness-AI-Trainer-With-Automatic-Exercise-Recognition-and-Counting

RiccardoRiccio/Fitness-AI-Trainer-With-Automatic-Exercise-Recognition-and-Counting

An extension of the previous 'Fitness-AI-Coach': a complete web application with real-time exercise recognition and counting. The exercise recognition model achieves 99% accuracy on the test set and 95% and 90% on two additional external datasets. || Give a star ⭐ to the reposito…

arxiv.org/abs/2011.12102v1

Do You Live a Healthy Life? Analyzing Lifestyle by Visual Life Logging

A healthy lifestyle is the key to better health and happiness and has a considerable effect on quality of life and disease prevention. Current lifelogging/egocentric datasets are not suitable for lifestyle analysis; consequently, there is no research...

arxiv.org/abs/1507.02989v1

A Bloom filter based semi-index on $q$-grams

We present a simple $q$-gram based semi-index, which allows to look for a pattern typically only in a small fraction of text blocks. Several space-time tradeoffs are presented. Experiments on Pizza & Chili datasets show that our solution is up to thr...

arxiv.org/abs/2006.03051v2

NewB: 200,000+ Sentences for Political Bias Detection

We present the Newspaper Bias Dataset (NewB), a text corpus of more than 200,000 sentences from eleven news sources regarding Donald Trump. While previous datasets have labeled sentences as either liberal or conservative, NewB covers the political vi...

arxiv.org/abs/2309.07376v2

VCD: A Video Conferencing Dataset for Video Compression

Commonly used datasets for evaluating video codecs are all very high quality and not representative of video typically used in video conferencing scenarios. We present the Video Conferencing Dataset (VCD) for evaluating video codecs for real-time com...