Skip to main content

Web to Markdown API

Web to Markdown API. Bring readable web content into your workflow.

Get the article content your text pipeline needs. Render the page, request Markdown and preserve its source for the next step.

01 / OVERVIEW

What is Web to Markdown API?

AdsCrawl Web to Markdown API extracts readable content from a browser-rendered page using POST /html with contentMode set to markdown. It returns a text representation for research, document processing and LLM workflows that need content rather than the complete page markup.

02 / USE CASES

Where it fits

01

Prepare text for retrieval

Use readable content as input to your own chunking and retrieval pipeline. Store the source URL alongside it so later answers can point back to the page.

02

Collect research material

Save article content in a portable text format for internal reference. Review the result before using it in summaries or reports.

03

Choose the right representation

Request Markdown for readable content, HTML for rendered markup, or named extraction fields when you need precise values. Avoid parsing a format your task does not need.

03 / QUICKSTART

Start with one complete workflow

  1. Create a server-side API key and choose the page URL.
  2. Send POST /html with contentMode set to markdown.
  3. Check the response status and save the text as page.md.
  4. Review extracted content and preserve its source URL in your downstream workflow.

Requires Bash and cURL. Set the ADSCRAWL_API_KEY environment variable on your server first.

Full request and response reference →
cURL
curl --fail-with-body -sS \
  -X POST "https://api.adscrawl.net/html" \
  -H "x-api-key: $ADSCRAWL_API_KEY" \
  -H "content-type: application/json" \
  -d '{
    "url": "https://www.adscrawl.net",
    "contentMode": "markdown",
    "waitUntil": "domcontentloaded"
  }' \
  --output page.md

04 / LIMITS & BILLING

Know the boundaries

Compare plans and allowances →
  • Readable-content extraction is not a lossless page copy. Navigation, interactive widgets and layout-specific elements may be omitted.
  • Markdown conversion does not verify the source's facts or grant permission to republish its content.
  • Dynamic content still depends on page loading and supported wait settings. Check the output rather than assuming every section was captured.
  • Requests consume credits and are subject to account limits. This endpoint does not supply a complete retrieval or embedding pipeline.

Keep API keys on the server. For HTTP 402 check your balance; for 429 reduce concurrency and retry with backoff.

05 / FAQ

Common questions

Is there a separate /markdown endpoint?

No. Use POST /html with contentMode set to markdown. The output mode selects Markdown within the existing endpoint.

Does Markdown include every element?

No. It is intended for readable content, not a complete representation of the visual layout or every interactive element.

Should I use this for exact product fields?

If you need a fixed set of named values, use the Web Scraping API with explicit field definitions. Markdown is useful when your workflow needs readable text.