Skip to content
Agents tracked: 157 Downloads (7d): 216M up 9.6% GitHub stars: 4.2M VS Code installs: 145M Releases (7d): 233 Agent pull requests (last week): 940K Updated Oct 6, 2026

AI agents for web scraping

Agents that open pages, click through them and return the data you asked for in a structured form.

By weekly package downloads, the leaders are Stagehand (2.5M), Browser Use (2.4M), Amazon Nova Act (71.7K). By GitHub stars, the leaders are Browser Use (117K), Stagehand (25.5K), Skyvern (23.1K). The fastest grower over the last 30 days is Stagehand, with downloads up 16%. As of Oct 6, 2026.

# Agent Category Pulse Downloads 7d 7d 30d VS Code installs Stars Latest release Price Last 90 days
14
Browser UseBrowser Use
Browser 60 2.4M up 16.3% — — 117K+88/day 0.13.1032d ago Free
20
StagehandBrowserbase
Browser 58 2.5M up 13.4% up 16.5% — 25.5K+15/day 3.7.339d ago Free + $20/mo
102
ClayClay
Sales 28 — — — — — — Free + $167/mo
104
SkyvernSkyvern AI
Browser 28 5,699 up 0.3% down 14.8% — 23.1K+8/day 1.0.555d ago Free + $29/mo
111
Amazon Nova ActAmazon (AWS)
Browser 23 71.7K down 3.5% down 41.0% — 918 3.4.187.05mo ago Usage-based
146
Perplexity ComputerPerplexity AI
General 0 — — — — — — —

Ranks are overall positions by Pulse Score. Click a column to sort by what matters to you. Methodology.

Agent or classic scraper

A classic scraper is fast and cheap but breaks when a page changes. A browser agent reads the page the way a person does, so it copes with layout changes and multi-step flows (search, filter, paginate) at a higher cost per page. Many teams use agents to build or repair extraction steps and classic code to run them at volume.

What each agent does for this task

Only agents whose own documentation describes this use are listed. Checked Oct 6, 2026.

Agent What it does Source
Browser Use Open-source browser agent that returns extracted web data as structured output docs
Stagehand extract() pulls structured data from web pages against a schema docs
Skyvern Browser automation that extracts structured data from websites against a schema docs
Amazon Nova Act act_get() extracts information from web pages, with an optional response schema docs
Clay Claygent researches and scrapes websites to enrich rows in a table docs
Perplexity Computer Automates browser actions and extracts data from websites docs

Stay within the rules

  • Read the site's terms and robots.txt. Many sites forbid automated collection; an agent does not change that.
  • Prefer an official API where one exists.
  • Do not collect personal data you have no lawful basis to process, and rate-limit what you run.

More options by job: the best AI agents in every category.