Back to catalog

Tavily · Search

Tavily Search and Extract API / Retrieve raw web content from specified URLs

Extract web page content from one or more specified URLs using Tavily Extract.

Private Gateway connection

Data below comes from the configured private Gateway. Provider activation remains governed by its evidence and policy gates.

unverifiedRequest shape unavailableUnverified
Published API evidence
Public metadata only. Response bodies, credentials, and internal review notes are never displayed.

No verification date is claimed. No response capture is claimed.

Request parameters

  • chunks_per_source body · optional

    Chunks are short content snippets (maximum 500 characters each) pulled directly from the source. Use `chunks_per_source` to define the maximum number of relevant chunks returned per source and to control the `raw_content` length. Chunks will appear in the `raw_content` field as: `<chunk 1> [...] <chunk 2> [...] <chunk 3>`. Available only when `query` is provided. Must be between 1 and 5.

  • extract_depth body · optional

    The depth of the extraction process. `advanced` extraction retrieves more data, including tables and embedded content, with higher success but may increase latency.`basic` extraction costs 1 credit per 5 successful URL extractions, while `advanced` extraction costs 2 credits per 5 successful URL extractions.

  • format body · optional

    The format of the extracted web page content. `markdown` returns content in markdown format. `text` returns plain text and may increase latency.

  • include_favicon body · optional

    Whether to include the favicon URL for each result.

  • include_images body · optional

    Include a list of images extracted from the URLs in the response. Default is false.

  • include_usage body · optional

    Whether to include credit usage information in the response. `NOTE:`The value may be 0 if the total successful URL extractions has not yet reached 5 calls. See our [Credits & Pricing documentation](https://docs.tavily.com/documentation/api-credits) for details.

  • query body · optional

    User intent for reranking extracted content chunks. When provided, chunks are reranked based on relevance to this query.

  • timeout body · optional

    Maximum time in seconds to wait for the URL extraction before timing out. Must be between 1.0 and 60.0 seconds. If not specified, default timeouts are applied based on extract_depth: 10 seconds for basic extraction and 30 seconds for advanced extraction.

  • urls body · required

    One or more URLs to extract content from (Tavily accepts up to 5 per request).

Sign in to run this operation, inspect live eligibility, and see governed execution and audit evidence. Sign in.

Tavily Search and Extract API / Retrieve raw web content from specified URLs · looot