Scraping, browsers and parsing for projects built with AI agents
Headless browsers, HTML parsing, crawling. 16 tools ranked by how many of 12,697 public projects built with AI coding agents use them, with downloads and GitHub stars. Updated Oct 10, 2026.
| Tool | Used in agent projects | Downloads / wk | GitHub stars | License |
|---|---|---|---|---|
| Beautiful Soup pypi Python HTML/XML parsing and screen-scraping library for web data extraction. |
160 (1.3%) | 69M | — | MIT License |
| Puppeteer npm Headless Chrome/Chromium automation for scraping, testing and screenshot generation. |
115 (0.9%) | 11.4M | 95.7K | Apache-2.0 |
| lxml pypi High-performance XML and HTML processing library combining libxml2/libxslt with Python. |
98 (0.8%) | 88.4M | 3,065 | BSD-3-Clause |
| Cheerio npm jQuery-like syntax for parsing and manipulating HTML/XML with server-side speed. |
94 (0.7%) | 26.6M | 30.5K | MIT |
| Playwright (Python) pypi High-level Python API for automating browser testing and web scraping. |
68 (0.5%) | 26M | 15K | — |
| fast-xml-parser npm High-performance XML parser and builder without native dependencies. |
58 (0.5%) | 83.3M | 3,139 | MIT |
| csv-parse npm Parse CSV files using Node.js streaming Transform API for memory-efficient processing. |
51 (0.4%) | 20.6M | 4,280 | MIT |
| Playwright Core npm Headless browser automation library for cross-browser testing and web scraping. |
49 (0.4%) | 122M | 97.4K | Apache-2.0 |
| soupsieve pypi CSS selector implementation for BeautifulSoup to query HTML and XML documents. |
39 (0.3%) | 70.4M | 271 | — |
| Selenium pypi Official Python bindings for browser automation and web scraping. |
34 (0.3%) | 7M | — | Apache-2.0 |
| jsonc-parser npm Parser for JSON with comments and trailing commas support. |
31 (0.2%) | 72.3M | 767 | MIT |
| Puppeteer Core npm Headless Chrome/Chromium control library for automation and testing. |
31 (0.2%) | 22.8M | 95.7K | Apache-2.0 |
| Bleach pypi Sanitizes HTML and prevents XSS attacks with a safelist-based approach. |
24 (0.2%) | 16.3M | 2,764 | Apache Software License |
| Parsimonious pypi Pure-Python PEG parser for building domain-specific language parsers. |
23 (0.2%) | 1.8M | 1,916 | MIT |
| RSS Parser npm Parse RSS feeds in Node.js and browser with simple, lightweight API. |
21 (0.2%) | 878K | 1,531 | MIT |
| feedparser pypi Parse RSS and Atom feeds with universal handling of feed formats. |
20 (0.2%) | 3.4M | 2,441 | BSD-2-Clause |
Nothing here matches. Search all tools →
Other categories
Web frameworks UI components and icons Styling and CSS Animation and motion 3D and graphics Charts and data visualisation Forms, validation and schemas State and data fetching Markdown, editors and content Desktop and mobile apps Utilities Model SDKs Agent, MCP and tool libraries Prompts, structured output and evals ML and data science Databases, ORMs and clients SDKs for hosted services Auth and security libraries Jobs, queues and workflows Files, images and media HTTP servers and APIs HTTP clients and networking Testing Linting and formatting Languages, types and runtimes Build tools and bundlers Dev workflow and config Logging, monitoring and analytics Crypto and web3 Game development