arxiv.org/abs/2401.02473v1
Recently, several works tackled the video editing task fostered by the success of large-scale text-to-image generative models. However, most of these methods holistically edit the frame using the text, exploiting the prior given by foundation diffusi...
github.com/LiheYoung/Depth-Anything
[CVPR 2024] Depth Anything: Unleashing the Power of Large-Scale Unlabeled Data. Foundation Model for Monocular Depth Estimation (⭐ 8020)
github.com/thomasdavis/backbonetutorials
As single page apps and large scale javascript applications become more prominent on the web, useful resources for those developers who are jumping the ship are crucial. (⭐ 2292)
github.com/corentinravoux/lelantos
Large-scale tomographic mapping using Lyman-alpha forest (⭐ 4)
github.com/devonfw-forge/devonfw4flutter-mts-app
Large-Scale Flutter Reference Application. An Extension of DevonFw's My Thai Star Project (⭐ 70)
arxiv.org/abs/2403.04701v4
Given the large-scale multi-modal training of recent vision-based models and their generalization capabilities, understanding the extent of their robustness is critical for their real-world deployment. In this work, we evaluate the resilience of curr...
arxiv.org/abs/2308.01566v2
An increasingly important building block of large scale machine learning systems is based on returning slates; an ordered lists of items given a query. Applications of this technology include: search, information retrieval and recommender systems. Wh...
arxiv.org/abs/2005.06580v2
Given that a MAC address can uniquely identify a person or a vehicle, continuous tracking over a large geographical scale has raised serious privacy concerns amongst governments and the general public. Prior work has demonstrated that simple hash-bas...
github.com/UCSC-VLAA/MedTrinity-25M
[ICLR 2025] This is the official repository of our paper "MedTrinity-25M: A Large-scale Multimodal Dataset with Multigranular Annotations for Medicine“ (⭐ 401)
github.com/devnen/Chatterbox-TTS-Server
Self-host the powerful Chatterbox TTS model. This server offers a user-friendly Web UI, flexible API endpoints (incl. OpenAI compatible), predefined voices, voice cloning, and large audiobook-scale text processing. Runs accelerated on NVIDIA (CUDA), AMD (ROCm)…
github.com/UMass-Embodied-AGI/TalkCuts
[NeurIPS 2025] TalkCuts: A Large-Scale Dataset for Multi-Shot Human Speech Video Generation (⭐ 28)
github.com/PaddlePaddle/PALM
a Fast, Flexible, Extensible and Easy-to-use NLP Large-scale Pretraining and Multi-task Learning Framework. (⭐ 185)
arxiv.org/abs/2308.02962v2
Science and Engineering fairs offer K-12 students opportunities to engage with authentic STEM practices. Particularly, students are given the chance to experience authentic and open inquiry processes, by defining which themes, questions and approache...
www.bing.com/ck/a?!&&p=236b9d4859b7abbae970be1f60a94db133a4757985f7a3be35be1927bc3da296JmltdHM9MTc3MjU4MjQwMA&ptn=3&ver=2&hsh=4&fclid=25fee722-35bf-6df4-0544-f03134c66cf4&u=a1aHR0cHM6Ly9naXRodWIuY29tL2RlZXBzZWVrLWFpL0RlZXBTZWVrLVIx&ntb=1
We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-Zero, a model trained via large-scale reinforcement learning (RL) without supervised fine-tuning …
arxiv.org/abs/astro-ph/0511791v1
Large scale simulations of Centaurs have yielded vast amounts of data, the analysis of which allows interesting but uncommon scenarios to be studied. One such rare phenomenon is the temporary capture of Centaurs as Trojans of the giant planets. Suc...
arxiv.org/abs/2308.11280v1
Hydrodynamic interactions can give rise to a collective motion of rotating particles. This, in turn, can lead to coherent fluid flows. Using large scale hydrodynamic simulations, we study the coupling between these two in spinner monolayers at weakly...
github.com/AIRI-Institute/nablaDFT
nablaDFT: Large-Scale Conformational Energy and Hamiltonian Prediction benchmark and dataset (⭐ 227)
arxiv.org/abs/2103.14294v2
Subgraph enumeration is a fundamental problem in graph analytics, which aims to find all instances of a given query graph on a large data graph. In this paper, we propose a system called HUGE to efficiently process subgraph enumeration at scale in th...
github.com/bcmi/Image-Harmonization-Dataset-iHarmony4
[CVPR 2020] The first large-scale public benchmark dataset for image harmonization. The code used in our paper "DoveNet: Deep Image Harmonization via Domain Verification", CVPR2020. Useful for image harmonization, image composition, etc. (⭐ 803)
github.com/PeiYiZhuo/zelenskyy-putin
After building a dataset by scraping 1796 press releases from the websites of the Kremlin and the President of Ukraine using the R package rvest, I conducted an analysis using the R packages tidyquant and tidytext that identified large differences between the…