arxiv.org/abs/2509.25027v1
Reinforcement learning has recently been explored to improve text-to-image generation, yet applying existing GRPO algorithms to autoregressive (AR) image models remains challenging. The instability of the training process easily disrupts the pretrain...
github.com/Graylog2/graylog2-images
Ready to run machine images (⭐ 240)
arxiv.org/abs/2511.16965v1
Synthesizing realistic cooked food images from raw inputs on edge devices is a challenging generative task, requiring models to capture complex changes in texture, color and structure during cooking. Existing image-to-image generation methods often p...
github.com/GoogleCloudPlatform/compute-archlinux-image-builder
A tool to build a Arch Linux Image for GCE (⭐ 286)
www.bing.com/ck/a?!&&p=264e1f0db41a7af8f8035f13a9fc9385f6c1855e65d5233fadabf7b0793f7fd0JmltdHM9MTc3MjQ5NjAwMA&ptn=3&ver=2&hsh=4&fclid=2e59e14d-483e-6efe-195b-f65f49ee6f25&u=a1aHR0cHM6Ly9zdXBwb3J0Lmdvb2dsZS5jb20vdHJhbnNsYXRlL2Fuc3dlci82MTQyNDgzP2hsPWZyJmNvPUdFTklFLlBsYXRmb3JtJTNEaU9T&ntb=1
Traduire du texte dans des images Dans l'application Traduction , vous pouvez traduire du texte à partir d'images enregistrées sur votre téléphone ou le texte détecté par votre appareil photo. Important : …
www.bing.com/ck/a?!&&p=8b9eb324ee880b566ba2a9e4eb18e515c65d9b6de6279c6011a481d757f2560cJmltdHM9MTc3MjQ5NjAwMA&ptn=3&ver=2&hsh=4&fclid=2e59e14d-483e-6efe-195b-f65f49ee6f25&u=a1aHR0cHM6Ly9zdXBwb3J0Lmdvb2dsZS5jb20vdHJhbnNsYXRlL2Fuc3dlci82MTQyNDgzP2hsPWZyLUZSJmNvPUdFTklFLlBsYXRmb3JtJTNERGVza3RvcA&ntb=1
Traduire du texte dans des images Google Traduction vous permet de traduire le texte qui figure dans des images depuis votre appareil. Important : L'exactitude de la traduction dépend de la clarté du …
arxiv.org/abs/2008.03426v1
Hyperspectral image (HSI) with high spectral resolution often suffers from low spatial resolution owing to the limitations of imaging sensors. Image fusion is an effective and economical way to enhance the spatial resolution of HSI, which combines HS...
www.bing.com/ck/a?!&&p=756dfcdb58290ce29909c489fc98065eeb0021cbcdec0801ff22b584e7106f73JmltdHM9MTc3MjQ5NjAwMA&ptn=3&ver=2&hsh=4&fclid=28a58dd4-4874-6e03-3104-9ac6494a6f7d&u=a1aHR0cHM6Ly93d3cuZ29vZ2xlLmNvbS5ici9pbWdocD9obD1wdC1QVA&ntb=1
Google Imagens. A pesquisa de imagens mais abrangente na Web.
arxiv.org/abs/2212.03239v2
Geometric camera calibration is often required for applications that understand the perspective of the image. We propose perspective fields as a representation that models the local perspective properties of an image. Perspective Fields contain per-p...
arxiv.org/abs/2506.21002v1
Scene text removal (STR) aims to erase textual elements from images. It was originally intended for removing privacy-sensitiveor undesired texts from natural scene images, but is now also appliedto typographic images. STR typically detects text regio...
github.com/facebookresearch/WSL-Images
Weakly Supervised Learning On Images (⭐ 602)
arxiv.org/abs/2512.08337v1
Generating BOLD images from T1w images offers a promising solution for recovering missing BOLD information and enabling downstream tasks when BOLD images are corrupted or unavailable. Motivated by this, we propose DINO-BOLDNet, a DINOv3-guided multi-...
arxiv.org/abs/2001.06265v1
Image-based virtual try-on for fashion has gained considerable attention recently. The task requires trying on a clothing item on a target model image. An efficient framework for this is composed of two stages: (1) warping (transforming) the try-on c...
arxiv.org/abs/2503.19902v2
The inherent ambiguity in defining visual concepts poses significant challenges for modern generative models, such as the diffusion-based Text-to-Image (T2I) models, in accurately learning concepts from a single image. Existing methods lack a systema...
github.com/webzhuce07/Digital-Image-Processing
Digital image processing is introduced in detail. all code are implementing by OpenCV 3.4 and C++ (⭐ 112)
en.wikipedia.org/wiki/Image-based_lighting
Image-based lighting (IBL) is a 3D rendering technique which involves capturing an omnidirectional representation of real-world light information as an
arxiv.org/abs/1411.6909v1
This paper proposes direct learning of image classification from user-supplied tags, without filtering. Each tag is supplied by the user who shared the image online. Enormous numbers of these tags are freely available online, and they give insight ab...
arxiv.org/abs/2403.15048v3
Leveraging large-scale Text-to-Image (TTI) models have become a common technique for generating exemplar or training dataset in the fields of image synthesis, video editing, 3D reconstruction. However, semantic structural visual hallucinations involv...
arxiv.org/abs/2312.01255v2
Diffusion-based image synthesis has attracted extensive attention recently. In particular, ControlNet that uses image-based prompts exhibits powerful capability in image tasks such as canny edge detection and generates images well aligned with these...
arxiv.org/abs/2306.02398v1
With the emergence of image super-resolution (SR) algorithm, how to blindly evaluate the quality of super-resolution images has become an urgent task. However, existing blind SR image quality assessment (IQA) metrics merely focus on visual characteri...