Apache Spark™ - Unified Engine for large-scale data analytics
Apache Spark is a multi-language engine for executing data engineering, data science, and machine learning on single-node machines or clusters.
Apache Spark is a multi-language engine for executing data engineering, data science, and machine learning on single-node machines or clusters.
Snowflake powers AI, data engineering, applications, and analytics on a trusted, scalable AI Data Cloud—eliminating silos and accelerating innovation.
Data Science algorithms and topics that you must know. (Newly Designed) Recommender Systems, Decision Trees, K-Means, LDA, RFM-Segmentation, XGBoost in Python, R, and Scala. (⭐ 134)
We introduce a new synthetic data generator PSP-HDRI$+$ that proves to be a superior pre-training alternative to ImageNet and other large-scale synthetic data counterparts. We demonstrate that pre-training with our synthetic data will yield a more ge...
This project analyzes Earthquake data from the USGS covering the past five years, focusing on seismic trends such as strongest quakes, shallow events, and regional activity. By combining insights from Geoscience and Seismology, the study highlights how earthqu…
This repo includes all the work which I have done during learning aws-data-engineer-certification-roadmap course (⭐ 0)
The FRED® App gets you the economic data you need—anytime, anywhere. Enjoy full access to over 840,000 economic data series from 118 regional, national, and international sources.
80 major categories of economic data. FRED: Download, graph, and track economic data.
Augmenting the training data of automatic speech recognition (ASR) systems with synthetic data generated by text-to-speech (TTS) or voice conversion (VC) has gained popularity in recent years. Several works have demonstrated improvements in ASR perfo...
QUERY(A2:E6,F2,FALSE) Syntax QUERY(data, query, [headers]) data - The range of cells to perform the query on. Each column of data can only hold boolean, numeric (including date/time types) or …
Data models related with Waste Water treatment (⭐ 11)
Automotive data including vehicle model, make, and year for database creation (⭐ 555)
This chapter presents the potential of interoperability and standardised data publication for cultural heritage resources, with a focus on community-driven approaches and web standards for usability. The Linked Open Usable Data (LOUD) design principl...
A sufficient level of data sovereignty is extremely difficult for consumers in practice. The EU General Data Protection Regulation guarantees comprehensive data subject rights, which must be implemented by responsible controllers through technical an...
J. Berberich, J. Köhler, M. A. Müller and F. Allgöwer, "Data-Driven Model Predictive Control With Stability and Robustness Guarantees," in IEEE Transactions on Automatic Control, vol. 66, no. 4, pp. 1702-1717, April 2021, doi: 10.1109/TAC.2020.3000182. (⭐ 82)
:blue_book: A Typescript companion to the book A Common-Sense Guide to Data Structures and Algorithms by Jay Wengrow (⭐ 30)
In this manuscript, I provide an updated interplanetary shock data base I published in previous works. This list has now 603 events. I also present and describe the data and methodologies used to compile this list. The main contribution of this work...
A dynamical analysis of the structure of the clusters of galaxies Abell 119 and Abell 133 is presented, using new redshift data combined with existing data from the literature. We compare our results with those from the X-ray data for these cluster...
The aim this study is discussed on the detection and correction of data containing the additive outlier (AO) on the model ARIMA (p, d, q). The process of detection and correction of data using an iterative procedure popularized by Box, Jenkins, and R...
specializes in integration platform as a service (iPaaS), data integration, API management, master data management, data preparation, and AI-powered