MCP servers for web scraping without a repository URL
Of the 346 MCP servers in the measured field of use Web scraping, the registry lists no repository URL for 121 – markedly more than the 78 the comparison group would lead one to expect.
121 servers 0.4% of 26,983 in the directory
This category in numbers
- 80 66 % Required secrets declared: no
- 91 75 % Icon formats (raw): nothing declared
- 104 86 % Schema version of the raw record (raw): https://static.modelcontextprotocol.io/schemas/2025-12-11/server.schema.json
- 111 92 % Execution location: remote
The most frequent value per property, counted across the entries of this category. The full distribution is further down.
Entries
Filter further
- Execution location: local
- Execution location: remote
- Required secrets declared: no
- Required secrets declared: yes
- recently updated first
Entries 1 to 50 of 121
| AI Crawler Index | dev.workers.pathwren.www/ai-crawler-index |
|---|---|
| AgentMart Procurement | online.agentmart/procurement |
| Amazon Reviews Scraper - Extract Product Data | com.ranovam/amazon-review-intelligence |
| Andi Search | com.andiai/andi-search |
| Ashby Job Scraper & API for Team Job Boards | com.ranovam/ashby-scraper |
| Atlassian Statuspage Scraper & JSON API Tool | com.ranovam/statuspage-scraper |
| Brainiall Web | com.brainiall/web |
| Crawler IP Verifier - real Googlebot or fake | dev.workers.pathwren.www/crawler-ip-verifier |
| Crawler Log Triage | dev.workers.pathwren.www/crawler-log-triage |
| Croma | com.usecroma/croma |
| Data Enrichment & Web Scraper API | io.github.Mila-Petrovich/data-enrichment |
| Data Quality Gate - deterministic post-scrape cleaner + verdict | io.github.aidatatools-dev/data-quality-gate |
| Discourse Forum Scraper — Topics & Posts | com.ranovam/discourse-forum-scraper |
| Domduck | com.domduck/domduck |
| Execute behance-portfolio-scraper-mcp | io.github.Evozim/behance-portfolio-scraper-mcp |
| EzBiz SEO & Marketing Analysis | com.ezbizservices/seo-marketing |
| FUTIA | io.futia/seo |
| Fastcrawl | net.fastcrawl/fastcrawl |
| Google Maps Business Contact Email Scraper | com.ranovam/maps-contact-cleaner |
| Google Maps Real Estate Agency Data Scraper | com.ranovam/realestate-agency-maps-scraper |
| Goro | ai.usegoro/goro |
| Greenhouse Jobs Scraper & API Data Extractor | com.ranovam/greenhouse-jobs-scraper |
| GroundAPI | io.github.qingkongzhiqian/groundapi |
| H/M Blindspot Challenge Platform | net.hogarmas/blindspot |
| Hydrafetch | com.hydrafetch/web |
| IDLE Protocol | com.earnidle/mcp |
| Instagram Hashtag Posts & Reels Scraper API | com.ranovam/instagram-hashtag-discovery |
| Instatus Status Page Scraper – JSON API Tool | com.ranovam/instatus-scraper |
| Iskat — Russian-language web search for agents | io.iskat/webdex |
| Kai AGI - Autonomous AI Agent | com.kai-agi.mcp/kai-agi |
| Kai AGI - Autonomous AI Agent | com.trycloudflare.annotated-lot-operational-tourism/kai-agi |
| Lever Jobs Scraper - Job Board Data Extractor | com.ranovam/lever-scraper |
| LexLint: compliance lint for AI, scraping, and privacy law | org.lexlint/lexlint |
| LinkedIn Lead Enrichment & Data Scraper API | com.ranovam/ai-lead-enricher-personalizer |
| LinkedIn Profile Data Scraper & Extractor Tool | com.ranovam/vertical-lead-pack |
| MESSORA | dev.messora/messora |
| Olostep MCP Server | io.github.olostep/olostep-mcp-server |
| Open Collective Expenses & Transactions Scraper | com.ranovam/opencollective-scraper |
| Pa1m SEO | io.github.retampweb/pa1m-seo |
| Perplexity API Platform | ai.perplexity/mcp-server |
| Perplexity API Platform (moved to ai.perplexity/mcp-server) | io.github.perplexityai/mcp-server |
| Personio Jobs Scraper - Clean JSON API Tool | com.ranovam/personio-scraper |
| Pinpoint Jobs Scraper - Job Board API Tool | com.ranovam/pinpoint-scraper |
| Prowl MCP | chat.prowl/prowl-mcp |
| Recruitee Jobs Scraper - Clean Job Data API | com.ranovam/recruitee-scraper |
| SE Ranking | com.seranking/mcp |
| SEEK Jobs Scraper: Australia & NZ Listings | com.ranovam/seek-scraper |
| ScrapeNest | dev.scrapenest/scrapenest |
| Scrappycoco | io.github.Albert-Tam/scrappycoco |
| SeaWeb | tech.seaweb/seaweb |
All 121 entries in the list, with filters and pagination
What this category selects
| Field of use, derived from the vendor description | Web scraping derived from: Description (raw) |
|---|---|
| Repository URL listed | no derived from: Repository (raw) |
A derived value carries its original data beside it; the derivation rule is published, with a version, on the methodology page.
Caveats, distribution and query
What this category does not say
- Field of use, derived from the vendor description is not declared but computed: from Description (raw). The rule is published, with a version, on the methodology page.
- Repository URL listed is not declared but computed: from Repository (raw). The rule is published, with a version, on the methodology page.
What is measured is what a manifest declares, not what a piece of software does. This registry fetches no repository URL, no endpoint and no package index; nothing here is verified. Which value comes from which source, and by which rule it was formed, is set out in the Methodology.
Every sentence here follows from the condition above, not from an assessment.
How this category is distributed
Required secrets declared
121| no | 80 66 % |
|---|---|
| yes | 41 34 % |
Icon formats (raw)
121| nothing declared | 91 75 % |
|---|---|
| image/png | 23 19 % |
| image/svg+xml | 5 4.1 % |
| image/png, image/svg+xml | 1 0.8 % |
| image/svg+xml, image/webp | 1 0.8 % |
Schema version of the raw record (raw)
121| https://static.modelcontextprotocol.io/schemas/2025-12-11/server.schema.json | 104 86 % |
|---|---|
| https://static.modelcontextprotocol.io/schemas/2025-09-29/server.schema.json | 15 12 % |
| https://static.modelcontextprotocol.io/schemas/2025-07-09/server.schema.json | 1 0.8 % |
| https://static.modelcontextprotocol.io/schemas/2025-10-17/server.schema.json | 1 0.8 % |
Execution location
121| remote | 111 92 % |
|---|---|
| local | 10 8.3 % |
Delivery form (raw)
121| remote | 111 92 % |
|---|---|
| package | 8 6.6 % |
| package and remote | 2 1.7 % |
Per property, the measured values inside this category. The share refers to the 121 entries of the category, not to the registry.
The same category through the interface
curl -s "https://api.tracevero.com/v1/eintraege?art=mcp_server&einsatzgebiet=websammlung&quelloffen_einsehbar=false&grenze=200"
The command carries exactly the condition of this page. 121 entries at 200 each are a single request; the total is in the X-Gesamt header.
Changes to these entries
https://tracevero.com/aenderungen.atom?thema=einsatz-websammlung-ohne-repository
Both addresses narrow the change stream to exactly the entries in this category. The feed carries an identifier of its own, so a feed reader keeps it apart from the full stream.
Categories sharing entries with this one
Per neighbour its own size and the counted number of shared entries. An category without a single shared entry is not listed here.
tracevero · https://tracevero.com/themen/einsatz-websammlung-ohne-repository