Extract web content
Extract raw web content from one URL or a batch of up to 20 URLs. Successful URLs appear in results and per-URL failures appear in failed_results; even when HTTP 200 returns an empty results array, check both arrays. Output order is not guaranteed. The endpoint supports query reranking, 1–5 chunks per source (only when query is provided), basic or advanced extraction, images, favicons, markdown or text format, and a 1–60 second timeout. Unknown request fields are ignored. Use Authorization: Bearer <key>, or pass api_key in the JSON body as a Tavily-compatibility fallback. Both BazaarLink base URLs are supported: https://api.bazaarlink.ai/v1 and https://bazaarlink.ai/api/v1. Every 5 successful URL extractions costs 1 credit with basic or 2 credits with advanced; failed URLs are free. The customer price is credits multiplied by the per-credit price set by the admin: see the BazaarLink pricing page and use usage.cost from the response. The free allowance shares the same daily credits as search; when it is exhausted with no balance, the response is 432. cURL: `curl https://api.bazaarlink.ai/v1/extract -H "Authorization: Bearer $BAZAARLINK_API_KEY" -H "Content-Type: application/json" -d '{"urls":["https://example.com/article"]}'` Python: `TavilyClient(api_key=..., api_base_url="https://api.bazaarlink.ai/v1").extract(urls=[...])` JavaScript: `tavily({ apiKey, apiBaseURL })` LangChain: `TavilyExtract(api_base_url=...)`
/v1/extractAuthorizations
AuthorizationrequiredAPI key as bearer token in the Authorization header.
Body
urlsrequiredURLs to extract. Pass one string or an array containing 1–20 URLs. An empty array or no valid URL returns 400.
["https://example.com/article"]queryUser intent used to rerank extracted content chunks.
"key findings"chunks_per_sourceMaximum relevant content chunks returned per source. Must be 1–5 and is available only when query is provided.
3extract_depthExtraction depth. advanced can retrieve tables and embedded content with higher latency; basic costs 1 credit per 5 successful URLs and advanced costs 2 credits.
basicadvanced"basic"include_imagesWhether to include image URLs extracted from each successful page.
trueinclude_faviconWhether to include the favicon URL in each successful result.
trueformatFormat of extracted content. markdown preserves Markdown structure; text returns plain text.
markdowntext"markdown"timeoutMaximum seconds to wait for extraction, from 1–60. When omitted, the default is 10 seconds for basic and 30 seconds for advanced.
30include_usageWhether to include credits and usage.cost in the response.
true