Agent♥︎Age
Catalog

ShadowCrawl

OfficialLive

by cortex-works · Rust

Stealth scraping & search. Bypasses Cloudflare, DataDome & LinkedIn via Cyborg HITL approach.

io.github.DevsHero/shadow-crawl (MCP Server)

Stealth scraping and search exposed as a Model Context Protocol (MCP) server. The description states it “bypasses Cloudflare, DataDome & LinkedIn” using a Cyborg Human-in-the-Loop (HITL) approach, and it supports automated web extraction and search for agent workloads.

🛠️ Key Features

  • Stealth scraping & search
  • Anti-bot handling via Cyborg HITL approach
  • Web extraction for AI agents
  • Token-efficient web retrieval (agent workload focus)
  • Built within an ecosystem component described as “Deep Research & Web Extraction”

🚀 Use Cases

  • Agent workflows needing reliable web retrieval
  • Environments with anti-bot protections (Cloudflare, DataDome, LinkedIn)
  • Optional Human-in-the-Loop fallback during extraction

⚡ Developer Benefits

  • MCP integration (“model-context-protocol” and “mcp” topics listed)
  • Emphasis on token-efficient retrieval for LLM tooling
  • Tooling-oriented stack tags: Rust, Playwright, VSCode, automated-testing

⚠️ Limitations

  • Specific anti-bot bypass targets are explicitly mentioned (Cloudflare, DataDome, LinkedIn), implying behavior is tied to these scenarios

Topics

ai-agentsrustsearxngself-hostedsemantic-searchweb-scrapinganti-bot-bypassdev-toolsknowledge-graphllm-toolsmodel-context-protocolstealthproxymcpmodel2vecsearchenginevscodeautomated-testingplaywright

Related servers

More in Search & Web