About ScrapeWhale

ScrapeWhale is a web data API that turns any brand’s traffic, ads, social profiles, search results, and web pages into structured JSON and clean markdown for teams building marketing AI agents.

What ScrapeWhale does

One-call brand dossier

Send a domain to /api/v1/brand and get its logo, estimated traffic, domain authority, Google ads, and homepage as markdown in one response. Each part succeeds or fails on its own, so one blocked source never sinks the report.

Traffic and SEO metrics

Look up estimated monthly visits, rankings, domain authority, organic footprint, and backlink counts for any domain. Use it to size a competitor or qualify a prospect before your agent writes a word.

Ad libraries

Pull the ads a brand runs from the Meta Ad Library and Google Ads, including creatives and advertiser details. Your agent sees the offers, calls to action, and copy a brand is paying to show.

Social and commerce data

Fetch profiles, posts, and followers from TikTok, Instagram, X, YouTube, LinkedIn, and more, plus TikTok Shop, Shopify, and Chrome Web Store listings. Every channel answers with the same key and response shape.

Search, maps, and trends

Run Google web searches, read AI Overviews, search Google Maps places, and track search interest over time, by region, or trending now. Agents get the same results a person would see, ready to reason over.

Any URL or file to markdown

Convert any public URL, or a PDF, Word, Excel, CSV, OpenDocument, Numbers, or image file up to 10 MB, into clean markdown. The output drops straight into a context window or a RAG pipeline.

What makes ScrapeWhale different

  1. 1

    JSON and markdown from the same call

    Every data endpoint accepts ?format=json, markdown, or both (the default). Structured data feeds your pipelines, clean markdown feeds your prompts, and both are stored and re-fetchable by id for free.

  2. 2

    Cache hits cost zero credits

    Public data is cached for 6 to 24 hours for most sources and 7 days for logos. If it was fetched recently, your request returns instantly with "cached": true and "creditsCharged": 0.

  3. 3

    Failed requests refund themselves

    A blocked source or empty page never costs a credit: failures return an error code and the credits go back automatically. The brand dossier only charges for the parts that succeeded.

  4. 4

    Credit packs that never expire

    There is no subscription. Most endpoints cost 1 credit, ads endpoints and AI Overview 2, and the full brand dossier up to 6. Packs are one-time purchases you spend whenever you like.

  5. 5

    Built for agents from the start

    One bearer key covers every endpoint, with no OAuth flow or per-platform credentials. Add the scrapewhale-mcp server to Claude or Cursor with one command, or point your agent at /llms.txt and /openapi.json.

Who uses ScrapeWhale

  • AI agent developers giving Claude, Cursor, or custom agents live brand data through MCP or REST.
  • Marketing and growth SaaS teams adding traffic, ads, and social data to their product without juggling several vendor keys.
  • Growth marketers and founders researching a competitor’s ads, traffic, and social presence from the Playground or free tools.
  • Agencies and consultants pulling a brand dossier before a pitch, an audit, or a strategy session.
  • Content and research teams turning web pages, PDFs, and YouTube transcripts into markdown for LLM context.
  • Ecommerce analysts tracking TikTok Shop products, Shopify catalogues, and Chrome Web Store listings.

How ScrapeWhale works

  1. 01

    Create an account

    Sign up and get 10 free credits with no card required. Create an API key in Settings.

  2. 02

    Call an endpoint

    Send your key as a bearer token over REST, run endpoints from the Playground, or connect the MCP server.

  3. 03

    Get JSON and markdown

    Choose the format per request. Every result is stored and can be re-fetched by id at no cost.

  4. 04

    Top up when you need to

    Buy a one-time credit pack when you run low. Failed calls are refunded and cache hits are free.

Questions or custom needs? Email us at support@scrapewhale.dev.

Key facts

Company name
ScrapeWhale
Type
Web data API (SaaS) for marketing AI agents
Founded
2026
Core offering
A REST API and MCP server that return brand traffic, ads, social profiles, search results, and any URL or file as structured JSON and clean markdown
Pricing
10 free credits on signup, no card required; one-time packs: Starter $9.90 (1,000 credits), Pro $19.90 (2,200 credits), Ultimate $49.90 (6,000 credits) (see pricing)
Contract terms
No subscription or contract; credit packs are one-time purchases that never expire
Services
REST API, MCP server, Playground, free tools, briefs, API reference (/docs/api, /llms.txt, /openapi.json)
Customers served
Builders of marketing AI agents, marketing and growth SaaS teams, marketers, agencies, and research teams

Frequently asked questions

How does pricing work?

ScrapeWhale uses credit packs, not a subscription. Most endpoints cost 1 credit, ads endpoints 2, and the full brand dossier up to 6. New accounts get 10 free credits without a card, packs never expire, failed requests are refunded automatically, and cache hits are free.

What does "cache hits are free" mean?

Public brand data is cached for 6 to 24 hours for most sources and 7 days for logos. If anyone fetched the same thing recently, your request returns instantly with "cached": true and costs zero credits.

What formats do I get back?

Every data endpoint returns structured JSON and clean markdown from the same call. Use ?format=json for pipelines, ?format=markdown for context windows, or both (the default). Both artifacts are stored and re-fetchable by id for free.

How do I use it from Claude or Cursor?

Run npx scrapewhale-mcp with your API key to add ScrapeWhale tools to Claude Desktop, Claude Code, Cursor, or any MCP client. There is also /llms.txt for prompt-level discovery and /openapi.json for codegen.

What can I convert to markdown?

Any public URL, including articles, documentation, and social posts, or an uploaded PDF, Word, Excel, CSV, OpenDocument, Numbers, or image file up to 10 MB. Either way you get clean, readable markdown with the title preserved.

Are there rate limits?

Each API key has a soft limit of 60 requests per minute as an abuse guard; credits are the real spend cap. If you need more, get in touch.

What data does ScrapeWhale collect?

Only public, logged-off data: the same pages anyone can open in a browser. Traffic figures are estimates, and you are responsible for respecting the terms of the sources you use.

Can I use the output commercially?

Yes. The JSON and markdown you generate are yours to use for any purpose, including commercial projects.