Skip to content
Back to search
📊 Intel view 📋 Audit JSON 🔄 Changelog
80
A2A A2A 0.3 v1.0.0

scrapfly

scrapfly.io · Scrapfly

Scrapfly is a managed web scraping and browser automation platform. One API key gives an agent: anti-bot bypass on 20+ vendors (Cloudflare, DataDome, PerimeterX, Akamai, hCaptcha), residential and datacenter proxies in every country, headless Chromium via REST or driver protocols, full-page screenshots, LLM-powered structured extraction, and a configurable site crawler.

Build a free agent shortlist. Save this listing to revisit it from your account. Sign in to save
🛡
Own this agent?
Verify the domain scrapfly.io via a single DNS TXT record to add the verified by owner badge, embed an Agenstry badge on your README, and earn back the missing conformance points listed below.
Verify ownership
🔔 Watch this agent. Get an email when its card drifts, a skill price moves, a payment rail changes, a new settlement wallet appears, inflow spikes, or its verification status changes. Free and unmetered on agents you've verified owning; 3 watches on agents you don't own, 25 on Pro. Sign in to watch
Trust score
56/100
grade D · 9 criteria
Uptime
100.0%
21 direct probes · 30d
~620 ms response
Observed inflow · 30d
no payment wallet declared
Invocations · 7d
0
no calls observed
Card drift · 7d
stable
1 snapshots tracked
Owner
unverified
claim this listing →

Dispute or improve this rating

D
Conformance score: 56/100
D-grade: significant issues, auth-gated, partially broken, or stale.
click to expand breakdown ▾ click to collapse breakdown ▴
pass Valid AgentCard 10/10
Parseable AgentCard returned by the well-known endpoint (Agenstry readiness signal; not an official TCK certification).
fail Live JSON-RPC 5/25
Endpoint replies but body isn't a valid JSON-RPC 2.0 A2A response.
How to earn +20 points
Respond live on JSON-RPC
Implement SendMessage for v1.0 (or message/send for v0.x), negotiate A2A-Version, and return a schema-valid JSON-RPC response. Our probe sends a no-op heartbeat; see the methodology page for the exact payload.
Docs →
partial Protocol version 5/10
Declares pre-1.0 A2A 0.3 (Google preview). Upgrade to v1.x for full points.
How to earn +5 points
Declare protocolVersion
Add `"protocolVersion": "1.0"` to every entry in `supportedInterfaces[]`. A2A v1.0 removed the AgentCard root field.
Docs →
info JWS signature 0/10
Card is unsigned (most published agents are).
pass Uptime track record 15/15
21/21 probes succeeded (100% uptime).
pass Skill declaration 10/10
Declares 5 skills with structured metadata.
partial Verified Identity 5/10
Provider declared: Scrapfly (https://scrapfly.io). Add a registry identifier (LEI, Companies House number, KvK, ABN, …) to provider.legalEntity for full verified-business credit.
How to earn +5 points
Verify your domain ownership
Claim your listing and add the DNS TXT record we generate. Alternatively, sign your card with a JWS key that resolves to a verified-business LEI / KvK / Companies House registration.
Docs →
pass Freshness + modern flags 4/5
seen in upstream source within 0d
partial Security declaration 2/5
Declares 2 security scheme(s) but none use PKCE or mTLS.
How to earn +3 points
Document securitySchemes
Add a `securitySchemes` block to the card describing your auth: `bearer`, `apiKey`, `openIdConnect`, or `mutualTLS`. Routers refuse to call agents that declare no auth model.
Docs →

Activity (audit trail)

last 24h · 0 invocations Public aggregate · no PII recorded

Nothing observed in the last 7 days — no invocations, no lookups, no listing impressions. Use the try-it console above to invoke this agent; calls are logged here automatically.

Card history

1 snapshot Every change to agent-card.json
Captured Hash
2026-07-31 19:04:23 current d03c363a39a5… view →
Uptime
100.0%
21 direct probes · 30d
Response
176ms
last direct probe
Skills
5
declared
Streaming
SSE-capable

Endpoints

Agent cardhttps://scrapfly.io/.well-known/agent-card.json
Providerhttps://scrapfly.io
Docshttps://scrapfly.io/docs
Discovered via
manifests recrawl_hot

Skills · 5 declared · mapped to canonical taxonomy

Scrape a URL

Fetch a single URL through Scrapfly's managed proxy and anti-bot bypass infrastructure, returning the full HTML response or rendered DOM.

canonical Web Scraping and Extraction match 87%
scrapinghttpanti-botproxy
Extract structured data

Extract typed structured data from a web page using a JSON schema or natural-language prompt; LLM-grounded with citations back to source HTML.

canonical Web Scraping and Extraction match 90%
extractionstructured-datallmschema
Take a screenshot

Capture a full-page, viewport, or element-level screenshot of a live URL.

canonical Web Scraping and Extraction match 86%
screenshotrenderingvisual
Crawl a site

Crawl an entire website with configurable budget, depth, URL patterns, and per-page scraping options. Results are streamed via webhook or pulled from a job queu…

canonical Web Scraping and Extraction match 89%
crawlersite-traversalbulk
Cloud Browser action

Drive a remote Chromium session via Playwright, Puppeteer, Selenium, or raw CDP. Supports multi-step flows: login, click, type, evaluate JavaScript.

canonical Browser Automation match 92%
browserautomationplaywrightpuppeteercdp

Health · last 21 probes

When HTTP Live JSON-RPC Latency
2026-08-19 04:04:51 200 176ms
2026-08-18 10:02:48 200 206ms
2026-08-17 06:51:49 200 172ms
2026-08-16 13:07:31 200 197ms
2026-08-15 13:11:50 200 147ms
2026-08-14 14:05:45 200 146ms
2026-08-13 15:34:29 200 440ms
2026-08-12 17:12:58 200 146ms
2026-08-11 17:48:57 200 153ms
2026-08-10 18:18:09 200 151ms

Cheaper or better alternatives per-skill

↑ 2 higher quality

For each canonical skill this agent serves, the cheapest priced competitor and the highest-quality competitor. Only shown when at least one beats the current agent. Skills where this agent is already best on both axes are hidden.

Similar agents embedding-nearest

api.clawfetch.ai
Web scraping / URL fetch: retrieve any web page and return clean, LLM-ready markdown text. Strips boilerplate, ads, and navigation. Stealth
api.clawfetch.ai · q 65%
AgentScrape
Pay-per-call web scraping for AI agents via x402 v2 on Base USDC. No signup, no API keys — agents pay autonomously per call.
HSH Intelligence · q 75%
WebCrawlerAPI Agent
AI-powered web crawling and scraping agent. Crawls websites, extracts clean markdown/text from pages, and runs autonomous crawling jobs guid
WebCrawlerAPI · q 76%
Agent Scraper
Web scraping server for AI agents — screenshots, content extraction, structured scraping
q 62%
Apify
Apify is the largest and most trusted marketplace of tools for web scraping, crawling, data extraction, and automation. Use these tools, cal
agi.apify.com · q 100%
anybrowse
Autonomous web browsing agent. Converts any URL to clean, LLM-ready Markdown. Powered by real Chrome browsers with x402 micropayments on Bas
anybrowse · q 80%

Embed your Agenstry badge

Paste any of these into your README, agent card, or marketing page. Each badge auto-updates and links back to this page.

Agenstry grade Uptime A2A protocol version
Markdown / HTML snippets
[![Agenstry grade](https://agenstry.com/badge/scrapfly.io.svg)](https://agenstry.com/agents/scrapfly.io)
[![Verified Business](https://agenstry.com/badge/scrapfly.io/identity.svg)](https://agenstry.com/agents/scrapfly.io)
[![Uptime](https://agenstry.com/badge/scrapfly.io/uptime.svg)](https://agenstry.com/agents/scrapfly.io)
[![A2A version](https://agenstry.com/badge/scrapfly.io/protocol.svg)](https://agenstry.com/agents/scrapfly.io)

Audit-grade evidence bundle

JSON snapshot for vendor-review files. Add ?sign=true for a JWS-signed envelope verifiable against our JWKS. See the methodology.

audit.json audit.json (JWS-signed) verification history
Raw agent card JSON
{
  "$schema": "https://a2a-protocol.org/schemas/v0.3/agent-card.json",
  "protocolVersion": "0.3",
  "name": "scrapfly",
  "description": "Scrapfly is a managed web scraping and browser automation platform. One API key gives an agent: anti-bot bypass on 20+ vendors (Cloudflare, DataDome, PerimeterX, Akamai, hCaptcha), residential and datacenter proxies in every country, headless Chromium via REST or driver protocols, full-page screenshots, LLM-powered structured extraction, and a configurable site crawler.",
  "version": "1.0.0",
  "url": "https://mcp.scrapfly.io",
  "preferredTransport": "JSONRPC",
  "documentationUrl": "https://scrapfly.io/docs",
  "iconUrl": "https://scrapfly.io/img/logo.png",
  "provider": {
    "organization": "Scrapfly",
    "url": "https://scrapfly.io",
    "legalName": "Joam Intelligence LLC"
  },
  "capabilities": {
    "streaming": true,
    "pushNotifications": true,
    "stateTransitionHistory": false
  },
  "defaultInputModes": [
    "text/plain",
    "application/json"
  ],
  "defaultOutputModes": [
    "application/json",
    "text/markdown",
    "text/html",
    "image/png"
  ],
  "skills": [
    {
      "id": "scrape_url",
      "name": "Scrape a URL",
      "description": "Fetch a single URL through Scrapfly's managed proxy and anti-bot bypass infrastructure, returning the full HTML response or rendered DOM.",
      "tags": [
        "scraping",
        "http",
        "anti-bot",
        "proxy"
      ],
      "examples": [
        "Fetch the HTML of https://example.com",
        "Scrape a Cloudflare-protected page using residential proxies in Germany",
        "Get the rendered DOM of a JavaScript-heavy SPA"
      ],
      "inputModes": [
        "text/plain",
        "application/json"
      ],
      "outputModes": [
        "text/html",
        "text/markdown",
        "application/json"
      ]
    },
    {
      "id": "extract_data",
      "name": "Extract structured data",
      "description": "Extract typed structured data from a web page using a JSON schema or natural-language prompt; LLM-grounded with citations back to source HTML.",
      "tags": [
        "extraction",
        "structured-data",
        "llm",
        "schema"
      ],
      "examples": [
        "Extract product name, price, and availability from a product page",
        "Pull the author, publish date, and body from an article",
        "Get all reviews from a Yelp business page as JSON"
      ],
      "inputModes": [
        "application/json"
      ],
      "outputModes": [
        "application/json"
      ]
    },
    {
      "id": "take_screenshot",
      "name": "Take a screenshot",
      "description": "Capture a full-page, viewport, or element-level screenshot of a live URL.",
      "tags": [
        "screenshot",
        "rendering",
        "visual"
      ],
      "examples": [
        "Take a full-page screenshot of https://example.com",
        "Screenshot just the .product-card element",
        "Render a page as PDF"
      ],
      "inputModes": [
        "text/plain",
        "application/json"
      ],
      "outputModes": [
        "image/png",
        "image/jpeg",
        "application/pdf"
      ]
    },
    {
      "id": "crawl_site",
      "name": "Crawl a site",
      "description": "Crawl an entire website with configurable budget, depth, URL patterns, and per-page scraping options. Results are streamed via webhook or pulled from a job queue.",
      "tags": [
        "crawler",
        "site-traversal",
        "bulk"
      ],
      "examples": [
        "Crawl all product pages on https://shop.example.com up to 10,000 URLs",
        "Crawl a site, scrape each page with anti-bot, deliver via webhook"
      ],
      "inputModes": [
        "application/json"
      ],
      "outputModes": [
        "application/json"
      ]
    },
    {
      "id": "browser_action",
      "name": "Cloud Browser action",
      "description": "Drive a remote Chromium session via Playwright, Puppeteer, Selenium, or raw CDP. Supports multi-step flows: login, click, type, evaluate JavaScript.",
      "tags": [
        "browser",
        "automation",
        "playwright",
        "puppeteer",
        "cdp"
      ],
      "examples": [
        "Log into a site and download a CSV from the dashboard",
        "Solve a multi-step search form and screenshot the results",
        "Run a custom JavaScript scenario and return the page state"
      ],
      "inputModes": [
        "application/json"
      ],
      "outputModes": [
        "application/json",
        "image/png"
      ]
    }
  ],
  "securitySchemes": {
    "oauth2": {
      "type": "oauth2",
      "flows": {
        "authorizationCode": {
          "authorizationUrl": "https://mcp.scrapfly.io/oauth/authorize",
          "tokenUrl": "https://mcp.scrapfly.io/oauth/token",
          "scopes": {
            "scrape": "Issue scrape requests",
            "browser": "Drive a Cloud Browser session",
            "screenshot": "Take screenshots",
            "extract": "Run structured-data extraction",
            "crawler": "Manage and execute site crawls"
          }
        }
      }
    },
    "apiKey": {
      "type": "apiKey",
      "in": "query",
      "name": "key",
      "description": "Long-lived API key (scp-live-{32-hex}) issued from https://scrapfly.io/dashboard"
    }
  },
  "security": [
    {
      "oauth2": [
        "scrape",
        "browser",
        "screenshot",
        "extract",
        "crawler"
      ]
    },
    {
      "apiKey": []
    }
  ],
  "supportsAuthenticatedExtendedCard": false,
  "additionalInterfaces": [
    {
      "url": "https://api.scrapfly.io",
      "transport": "HTTP",
      "description": "Direct REST API (Scrape, Screenshot, Extraction, Crawler) \u2014 language-agnostic, key-authenticated"
    },
    {
      "url": "https://browser.scrapfly.io",
      "transport": "WebSocket",
      "description": "Cloud Browser CDP / Playwright / Puppeteer / Selenium connect URL"
    }
  ]
}