arxiv.org/abs/1404.3466v1
A well-known problem in numerical ecology is how to recombine presence-absence matrices without altering row and column totals. A few solutions have been proposed, but all of them present some issues in terms of statistical robustness (i.e. their cap...
arxiv.org/abs/2511.07306v1
An attorney submitted a 'right to be forgotten' delisting request to Google, regarding a blog post about a criminal conviction of the attorney in another country. The Rotterdam District Court ruled that Google may no longer link to the blog post when...
arxiv.org/abs/2407.05735v3
The ICRA conference is celebrating its $40^{th}$ anniversary in Rotterdam in September 2024, with as highlight the Happy Birthday ICRA Party at the iconic Holland America Line Cruise Terminal. One month later the IROS conference will take place, whic...
arxiv.org/abs/2410.17880v1
Residential location choices are traditionally modelled using factors related to accessibility and socioeconomic environments, neglecting the importance of local street-level conditions. Arguably, this neglect is due to data practices. Today, however...
arxiv.org/abs/2411.10811v1
The study aimed at detecting cartel collusion involved analyzing decisions of the Russian Federal Antimonopoly Service and data on auctions. As a result, a machine learning model was developed that predicts with 91% accuracy the signs of collusion be...
www.bing.com/ck/a?!&&p=c39dcf4fe23b62ace4e05ceb7e78d9a712231ac0849e8fa2317441e0043c5606JmltdHM9MTc3MjQwOTYwMA&ptn=3&ver=2&hsh=4&fclid=0a9cc673-c647-6b7e-004d-d163c7156a55&u=a1aHR0cHM6Ly93d3cuYXV0b2Rlc2suY29tL3Byb2R1Y3RzL2ZhYnJpY2F0aW9uL292ZXJ2aWV3P21zb2NraWQ9MGE5Y2M2NzNjNjQ3NmI3ZTAwNGRkMTYzYzcxNTZhNTU&ntb=1
Use a common library of component, material, and part data in ESTmep, CADmep, and CAMduct to quickly generate accurate, detailed, and constructible fabrication designs.
arxiv.org/abs/2106.06524v1
Wax is what you put on a surfboard to avoid slipping. It is an essential tool to go surfing... We introduce WAX-ML a research-oriented Python library providing tools to design powerful machine learning algorithms and feedback loops working on streami...
github.com/MayCooper/Snowflake-SQL-WalmartCommerceDB-AWS-Project
Creating a Walmart Commerce DB utilizing AWS, Snowflake, SnowSQL CLI, Apache Arrow, Python, ingesting Parquet & JSON data, and creating Materialized Views, Search Optimization, Transient Tables, Clustering, Access Control Roles, monitoring and analyzing query…
github.com/jwebber1/WalmartComparison
Compares items from Walmart using an api (https://developer.walmartlabs.com), a database (sqlite3), and an xml reading java program. (⭐ 0)
www.reddit.com/r/dataisbeautiful/comments/1oxqgds/oc_nutrient_density_of_highprotein_foods/
...
arxiv.org/abs/2312.16511v1
Supplying data augmentation to conversational question answering (CQA) can effectively improve model performance. However, there is less improvement from single-turn datasets in CQA due to the distribution gap between single-turn and multi-turn datas...
arxiv.org/abs/2105.12887v1
Recently there has been a huge interest in dialog systems. This interest has also been developed in the field of the medical domain where researchers are focusing on building a dialog system in the medical domain. This research is focused on the mult...
github.com/magicalpanda/MagicalRecord
Super Awesome Easy Fetching for Core Data! (⭐ 10740)
www.bing.com/ck/a?!&&p=a3aa94d76585d934c564fa51e030036c4620fc722e3f4ae19a02c575eadfa296JmltdHM9MTc3MjQwOTYwMA&ptn=3&ver=2&hsh=4&fclid=3578a85e-dd5c-6f5b-0911-bf4edc7d6efe&u=a1aHR0cHM6Ly93d3cud2VhdGhlcnRvbW9ycm93Lm5ldC8&ntb=1
4 days ago · Weather Tomorrow Offers a 15 day long-range forecast or an hour by hour forecast for the current day. Data is available for major cities of the world.
github.com/jmzeng1314/NGS-pipeline
By study this, it won't be costly or time-consuming to customize a NGS data analysis pipeline (⭐ 330)
arxiv.org/abs/2407.03036v1
Handling distribution shifts from training data, known as out-of-distribution (OOD) generalization, poses a significant challenge in the field of machine learning. While a pre-trained vision-language model like CLIP has demonstrated remarkable zero-s...
arxiv.org/abs/2410.14758v2
By embedding discrete representations into a continuous latent space, we can leverage continuous-space latent diffusion models to handle generative modeling of discrete data. However, despite their initial success, most latent diffusion methods rely...
arxiv.org/abs/2303.03717v1
Self-supervised learning (SSL) has recently shown remarkable results in closing the gap between supervised and unsupervised learning. The idea is to learn robust features that are invariant to distortions of the input data. Despite its success, this...
arxiv.org/abs/2001.11739v3
Intrinsic dimensionality (ID) is one of the most fundamental characteristics of multi-dimensional data point clouds. Knowing ID is crucial to choose the appropriate machine learning approach as well as to understand its behavior and validate it. ID c...
arxiv.org/abs/2304.11663v1
Deep equilibrium models (DEQs) have proven to be very powerful for learning data representations. The idea is to replace traditional (explicit) feedforward neural networks with an implicit fixed-point equation, which allows to decouple the forward an...