181 results for Crawler

arxiv.org/abs/2602.20705v1

The Careless Coupon Collector's Problem

We initiate the study of the Careless Coupon Collector's Problem (CCCP), a novel variation of the classical coupon collector, that we envision as a model for information systems such as web crawlers, dynamic caches, and fault-resilient networks. In C...

github.com/program-spiritual/KongFuOfArchitect

program-spiritual/KongFuOfArchitect

(Updating!) Architect's Kung Fu tutorial collection Article collection contains paradigm programming microservices essential algorithms Security attack Assembly Crawler Reverse penetration test...etc. (⭐ 364)

arxiv.org/abs/1601.06919v1

BUbiNG: Massive Crawling for the Masses

Although web crawlers have been around for twenty years by now, there is virtually no freely available, opensource crawling software that guarantees high throughput, overcomes the limits of single-machine systems and at the same time scales linearly...

github.com/vinta/haul

vinta/haul

An Extensible Image Crawler (⭐ 162)

github.com/zubair-trabzada/geo-seo-claude

zubair-trabzada/geo-seo-claude

GEO-first SEO skill for Claude Code. Comprehensive AI search optimization for any website — citability scoring, AI crawler analysis, brand authority, schema markup, platform-specific optimization, and PDF reports. (⭐ 146)

github.com/vinigracindo/pyfutebol

vinigracindo/pyfutebol

Simples crawler para obter resultados dos jogos de futebol (⭐ 48)

github.com/bernard0047/Profile-Exposer

bernard0047/Profile-Exposer

First place solution for ThaparWings hackathon; an intelligent crawler which collects profiles of important profiles present in deep links by navigating from given set of root URL(s). (⭐ 8)

github.com/StanGirard/seo-audits-toolkit

StanGirard/seo-audits-toolkit

SEO & Security Audit for Websites. Lighthouse & Security Headers crawler, Sitemap/Keywords/Images Extractor, Summarizer, etc ... (⭐ 776)

github.com/unclecode/crawl4ai

unclecode/crawl4ai

?? Crawl4AI: Open-source LLM Friendly Web Crawler & Scraper. Don't be shy, join here: https://discord.gg/jP8KfhDhyN (⭐ 61495)

github.com/code4craft/webmagic

code4craft/webmagic

A scalable web crawler framework for Java. (⭐ 11700)

www.reddit.com/r/cryptids/comments/15gr6rr/crawlers_what_are_they/

Crawlers? What are they?

No, I'm not talking about the Fresno Night Crawlers. I posted an incident about what people have told me is a crawler based on description. Tall, gangly, on all fours, very pale with eye shine and...

arxiv.org/abs/2308.04689v2

Web crawler strategies for web pages under robot.txt restriction

In the present time, all know about World Wide Web and work over the Internet daily. In this paper, we introduce the search engines working for keywords that are entered by users to find something. The search engine uses different search algorithms f...

en.wikipedia.org/wiki/Matt_Dinniman

Matt Dinniman - Wikipedia

American author known for the science fantasy LitRPG book series Dungeon Crawler Carl. Dinniman began writing as a child and started his first novel while