Prepare text for retrieval
Use readable content as input to your own chunking and retrieval pipeline. Store the source URL alongside it so later answers can point back to the page.
Web to Markdown API
Get the article content your text pipeline needs. Render the page, request Markdown and preserve its source for the next step.
01 / OVERVIEW
AdsCrawl Web to Markdown API extracts readable content from a browser-rendered page using POST /html with contentMode set to markdown. It returns a text representation for research, document processing and LLM workflows that need content rather than the complete page markup.
02 / USE CASES
Use readable content as input to your own chunking and retrieval pipeline. Store the source URL alongside it so later answers can point back to the page.
Save article content in a portable text format for internal reference. Review the result before using it in summaries or reports.
Request Markdown for readable content, HTML for rendered markup, or named extraction fields when you need precise values. Avoid parsing a format your task does not need.
03 / QUICKSTART
Requires Bash and cURL. Set the ADSCRAWL_API_KEY environment variable on your server first.
Full request and response reference →curl --fail-with-body -sS \
-X POST "https://api.adscrawl.net/html" \
-H "x-api-key: $ADSCRAWL_API_KEY" \
-H "content-type: application/json" \
-d '{
"url": "https://www.adscrawl.net",
"contentMode": "markdown",
"waitUntil": "domcontentloaded"
}' \
--output page.mdKeep API keys on the server. For HTTP 402 check your balance; for 429 reduce concurrency and retry with backoff.
05 / FAQ
No. Use POST /html with contentMode set to markdown. The output mode selects Markdown within the existing endpoint.
No. It is intended for readable content, not a complete representation of the visual layout or every interactive element.
If you need a fixed set of named values, use the Web Scraping API with explicit field definitions. Markdown is useful when your workflow needs readable text.