← TavilyCONTENT HISTORY

Update to Tavily

Snapshot Sep 30, 2026 · 22:45 UTC · version 2.0.0

Collection source: not recorded for this historical snapshot.

WHAT CHANGED · RULE-BASED ANALYSIS

First saved snapshot

No earlier snapshot is available to establish a change.

Compare saved observations

Download comparison JSON
Full technical diff · 0 changed fields
Full snapshot data
{
  "name": "tavily-extract",
  "description": "Extract clean markdown or text content from specific URLs via the Tavily CLI. Use this skill when the user has one or more URLs and wants their content, says \"extract\", \"grab the content from\", \"pull the text from\", \"get the page at\", \"read this webpage\", or needs clean text from web pages. Handles JavaScript-rendered pages, returns LLM-optimized markdown, and supports query-focused chunking for targeted extraction. Can process up to 20 URLs in a single call.\n",
  "included_files": [],
  "skill_md_contents": "---\nname: tavily-extract\ndescription: |\n  Extract clean markdown or text content from specific URLs via the Tavily CLI. Use this skill when the user has one or more URLs and wants their content, says \"extract\", \"grab the content from\", \"pull the text from\", \"get the page at\", \"read this webpage\", or needs clean text from web pages. Handles JavaScript-rendered pages, returns LLM-optimized markdown, and supports query-focused chunking for targeted extraction. Can process up to 20 URLs in a single call.\nallowed-tools: Bash(tvly *)\n---\n\n# tavily extract\n\nExtract clean markdown or text content from one or more URLs.\n\n## Before running\n\nRun extract directly when `tvly` is available. Extract supports capped keyless\naccess, so do not look for an API key or authenticate before the first request.\n\nIf `tvly` is missing, follow the [tavily-cli setup](../tavily-cli/SKILL.md#setup)\nbefore retrying. If the keyless cap is reached in an interactive session, run\n`tvly login` to open browser OAuth, then retry the original extraction once. In\nan unattended environment, report the cap and authentication options instead\nof starting an interactive flow. Do not start a second login immediately after\nguided setup has completed.\n\n## When to use\n\n- You have a specific URL and want its content\n- You need text from JavaScript-rendered pages\n- Step 2 in the [workflow](../tavily-cli/SKILL.md): search → **extract** → map → crawl → research\n\n## Quick start\n\n```bash\n# Single URL\ntvly extract \"https://example.com/article\" --json\n\n# Multiple URLs\ntvly extract \"https://example.com/page1\" \"https://example.com/page2\" --json\n\n# Query-focused extraction (returns relevant chunks only)\ntvly extract \"https://example.com/docs\" --query \"authentication API\" --chunks-per-source 3 --json\n\n# JS-heavy pages\ntvly extract \"https://app.example.com\" --extract-depth advanced --json\n\n# Save to file\ntvly extract \"https://example.com/article\" -o article.json\n```\n\n## Options\n\n| Option | Description |\n|--------|-------------|\n| `--query` | Rerank chunks by relevance to this query |\n| `--chunks-per-source` | Chunks per URL (1-5, requires `--query`) |\n| `--extract-depth` | `basic` (default) or `advanced` (for JS pages) |\n| `--format` | `markdown` (default) or `text` |\n| `--include-images` | Include image URLs |\n| `--timeout` | Max wait time (1-60 seconds) |\n| `-o, --output` | Save the JSON response to a file |\n| `--json` | Structured JSON output |\n\n## Extract depth\n\n| Depth | When to use |\n|-------|-------------|\n| `basic` | Simple pages, fast — try this first |\n| `advanced` | JS-rendered SPAs, dynamic content, tables |\n\n## Tips\n\n- **Max 20 URLs per request** — batch larger lists into multiple calls.\n- **Use `--query` + `--chunks-per-source`** to get only relevant content instead of full pages.\n- **Try `basic` first**, fall back to `advanced` if content is missing.\n- **Set `--timeout`** for slow pages (up to 60s).\n- **Inspect `failed_results` even after exit code 0.** A successful request can\n  still return no extracted pages. Retry the affected URL with `advanced` when\n  appropriate, otherwise report the per-URL failure instead of treating the\n  request as complete.\n- If search results already contain the content you need (via `--include-raw-content`), skip the extract step.\n\n## See also\n\n- [tavily-search](../tavily-search/SKILL.md) — find pages when you don't have a URL\n- [tavily-crawl](../tavily-crawl/SKILL.md) — extract content from many pages on a site\n"
}

SHA-256: d48e230cbd7e5848bdb1302794756592676dbcb45da96e7fefdd616f5c68e7a8