NewMCP server live

Web Scraping API, LLM-ready out of the box.

Get Markdown, HTML, JSON or plain text from any webpage, including JavaScript ones. Rendering, CAPTCHA solving, and bot detection are handled for you, starting at 1 credit a request.

4.9average across
G2 — 4.8 out of 5Google — 4.9 out of 5Product Hunt — 5.0 out of 5
  • 500 free credits every month
  • No card to start
  • Failed requests cost nothing
Playground
format
Request
POST /webscraping
{
  "url": "stripe.com",
  "format": [
    "markdown-llm"
  ]
}
loading example…
Want every endpoint and every parameter?
The Geekflare playground has every API and option, plus saved requests. Free with 500 credits a month.
Open the playground
Trusted by teams at
PfizerNBCUniversalTCSHostingerKissflowLookoutPlivoClearSaleSparkianCBSplitOmreon
PfizerNBCUniversalTCSHostingerKissflowLookoutPlivoClearSaleSparkianCBSplitOmreon
Handled for you

The six problems you don't have to solve.

Each is a system you would otherwise have to build, run, and maintain. The first five are included with every request at 1 credit.

JavaScript rendering

Pages are fetched without a browser first, then rendered in headless Chrome only when needed. React, Vue, and Angular applications return their rendered content, while static pages stay fast.

renderJS

CAPTCHA solving

Common CAPTCHA challenges are handled in the request path. There's no separate solver to integrate and no extra charge.

included

Anti-bot bypass

Browser fingerprints are managed on our side to handle common anti-bot systems. Stealth mode reduces browser automation signals for stricter sites.

stealth: true

Seven output formats

Choose from Markdown, HTML, or plain text, each with an LLM variant, plus structured JSON. Request multiple formats in a single call.

format: []

Structured extraction

Extract structured fields with CSS or XPath schemas, or use ready-made product and contact templates to get typed JSON.

extractionMode

Geo-targeted proxies

Route requests through a proxy in any of 194 countries to see localized pricing, content, and availability. Geo-targeting is opt-in and adds 4 credits when used.

proxyCountry
Geo-targeting

See the page the way locals see it.

Some sites block foreign traffic, and many show different prices and content by country. Add a country code and the request goes through a proxy, so you get the local page and don't get blocked.

  • proxyCountry routes a request through any of 194 countries
  • proxyMode: "auto" tries direct first and uses a proxy only if the site blocks you
  • Adds 4 credits only on requests that use a proxy
0
Countries
Country
Level of geo-targeting
Opt-in
Only when you need
POST/webscraping{ "url": "shop.com", "proxyCountry": "us" }from United States
Parameters

The parameters you'll use most.

See what each parameter does, what it costs, and when to use it without digging through the API reference.

ParameterTypeWhat it doesCredits
urlstringThe page to fetch. Any public URL.—
formatstring[]markdown, markdown-llm, html, html-llm, text, text-llm or json. Ask for several at once.included
renderJSbooleanRendering is automatic. Set true or false to force headless Chrome on or off.included
stealthbooleanRemoves webdriver signals and patches navigator properties for sites with stricter bot checks.included
waitTimenumberSeconds to wait after load, for lazy-loaded content.included
devicestringdesktop or mobile.included
extractionModestringWith format json: cssSchema, xpathSchema, or template with product or contact.included
fileOutputbooleanReturn a link to a file with the result instead of the content inline. Handy for large pages. Defaults to false.included
proxyCountrystringRoute the request through a proxy in this country, e.g. "de".+4
proxyModestringauto tries without a proxy first and uses one only if the site blocks the request.+4 if used
aiPromptobjectAsk AI to extract or analyze the page. Runs on the page's Markdown.+6

These are the parameters you'll use most. The endpoint reference has the complete parameter list, response schema, and error codes.

Endpoint reference
Use cases

What people actually use it for.

Clean Markdown straight into a vector store

Feed webpages into a retrieval pipeline without building your own extraction layer. Our LLM-ready format removes navigation, footers, cookie banners, and other page clutter, leaving the content you actually want to embed.

  • Headings are preserved for meaningful chunk boundaries
  • Navigation, footers and cookie banners removed
  • Multiple formats from a single request when you need both
Typical request
POST /webscraping
{
  "url": "https://techcrunch.com/category/artificial-intelligence/",
  "format": ["markdown-llm"]
}
What comes back
headingskept
# and ## structure intact
nav, footer, cookie bannerremoved
links and tableskept as Markdown
Quickstart

Your first scraping, in a few lines.

quickstart.pyofficial SDK
# pip install geekflare-api
from geekflare_api.client import GeekflareClient
from geekflare_api.models import WebScrapeDto

with GeekflareClient(api_key="<api-key>") as client:
    result = client.web_scrape(
        WebScrapeDto(
            url="https://example.com",
            format=["markdown"]
        )
    )
    print(result)
Ways in

One API. Every stack.

The same key works from an agent, a workflow builder, an SDK or your own service. Pick the lane you already live in.

MCP server

One remote endpoint exposes Geekflare's tools to any MCP client. Your agent discovers them itself: no glue code, no per-tool wiring.

ClaudeCursorCodexany MCP client
Set up the MCP server

LLM & agent pipelines

LLM-ready Markdown straight into a retrieval pipeline, with the boilerplate stripped so you spend context on content, not navigation.

RAGagentsmarkdown-llmgrounded search
See the solutions

Workflow builders

Run any endpoint as a step in a visual scenario. Useful when the person who needs the data does not write code.

ZapierMaken8n
Browse integrations

SDKs & REST

First-party typed SDKs for Python and Node. Everything else talks plain REST, from any language.

PythonNode.jsPostman
Open the API reference
Pricing

One credit a page. Here is what that buys.

Credits are shared with every other Geekflare endpoint. Failed requests cost nothing.

5002M

Growth covers up to ~100,000 pages a month (100K credits), about $0.69 per 1,000 pages. Or $58/mo billed yearly.

pages
60,000
Plan
Growth · $69/mo
FreeNo card
$0/mo
~500
pages / month
  • 500 credits / month
  • 1 team member
  • 7 days log retention
  • 1 request per second
Starter
$19/mo
~10,000
pages / month
  • 10K credits / month
  • 3 team members
  • 30 days log retention
  • 5 requests per second
GrowthMost popular
$69/mo
~100,000
pages / month
  • 100K credits / month
  • 5 team members
  • 30 days log retention
  • 10 requests per second
Business
$349/mo
~1,000,000
pages / month
  • 1M credits / month
  • 25 team members
  • 90 days log retention
  • 25 requests per second
Start freeCompare plans and credit packsEvery account starts free. Upgrade from the dashboard when you need more, or buy a credit pack from $10.
TESTIMONIALS

Teams that stopped maintaining scrapers.

Read all reviews
“Found Geekflare API to get markdown from URL for my AI agents. It is fast and cheaper and works on almost every website.”
Ram DasiArchitect, PA Consulting
“After trying many scraping services, I selected Geekflare to scrape public directories. Mainly for two reasons - it is cheaper and fast.”
Vishu SharmaCTO
“Documentation was easy to follow. We were scraping dynamic pages within hours. Very reliable service.”
Marco SilvaAnalytics Engineer
Questions

Before you write the integration.

Everything else is in the endpoint docs, and our support team replies to every email.

Open the endpoint docs

1 credit per request, whether or not JavaScript is rendered. CAPTCHA solving and anti-bot bypass are included at that price. Routing through a proxy adds 4 credits (5 total), and an aiPrompt adds 6. Failed requests cost nothing.

Yes. Pages that need JavaScript are rendered in headless Chrome, so React, Vue and Angular apps return the content a browser would see. Static pages are fetched without a browser, which is faster. Set renderJS to force either way.

Yes. Add proxyCountry with a country code, for example "de", to route the request through a proxy in that country and get localized pricing and content. Proxies cover 194 countries at country level and add 4 credits per request.

markdown-llm for LLM and RAG pipelines, json when you want typed fields, html when you have your own parser downstream, and text for search indexing or NLP. You can request more than one format in a single call.

Yes. With format: ["json"], use a CSS or XPath schema to pick fields, or set extractionMode to template with product or contact for ready-made extraction. For free-form questions about a page, aiPrompt runs AI over its Markdown.

Common CAPTCHA types are solved automatically, and fingerprints are managed to get past Cloudflare-style bot detection. For stricter sites, stealth: true removes webdriver signals.

No. You can run requests in the dashboard playground, or connect the API to Zapier, Make or n8n. AI agents can call it through our MCP server.

Scraping publicly available data is generally legal, but it depends on the site's terms of service and on privacy laws such as GDPR and CCPA. It is your responsibility to target public data, respect robots.txt and avoid collecting personal information.

Web Scraping returns a page's content as Markdown, HTML, JSON or text. Meta Scraping returns only its metadata: title, description, Open Graph, Twitter cards and Schema.org JSON.

500 pages, no card, one curl away.

The free tier renews every month. Scrape one of the pages and see what comes back.