Bleach
Sanitizes HTML and prevents XSS attacks with a safelist-based approach.
Scraping, browsers and parsing PyPI Apache Software License
Used in agent projects
24
0.2% of 12,697 public repos built with AI coding agents
Rank in scraping, browsers and parsing
#13 of 16
by use in those projects
Downloads
16.3M
PyPI, last 7 days, bleach
GitHub stars
2,764
last push Jun 5, 2026
How Bleach is used
Bleach is a dependency in 24 of the 12,697 public GitHub projects set up for AI coding agents that we scanned (0.2%), #13 of 16 in scraping, browsers and parsing. The most used alternative there is Beautiful Soup. Projects that use it most often also use Requests (83.3%), urllib3 (66.7%), jsonschema (62.5%).
Often used with Bleach
| Requests | 83.3% of Bleach projects · 15.7× more than average |
|---|---|
| urllib3 | 66.7% of Bleach projects · 43.9× more than average |
| jsonschema | 62.5% of Bleach projects · 61.5× more than average |
| pandas | 62.5% of Bleach projects · 20.7× more than average |
| NumPy | 62.5% of Bleach projects · 16.3× more than average |
| Beautiful Soup | 58.3% of Bleach projects · 46.3× more than average |
| defusedxml | 58.3% of Bleach projects · 308.6× more than average |
| Jinja2 | 58.3% of Bleach projects · 42.3× more than average |
By coding agent
| Claude Code | 0.3% of projects set up for it |
|---|---|
| GitHub Copilot | 0.2% of projects set up for it |
| Cursor | 0.2% of projects set up for it |
Share of public repos configured for each agent that list Bleach as a dependency.
Install
| Package | pip install bleach |
|---|---|
| Main stacks | Python (23), Next.js (1) |
| Example projects | FeatureFactory-io/mimir, AndreAugusto11/XChainDataGen, nihilistau/CosySim, brandoz2255/Harvis, gabrielfior/solana-analytics-101 |
Alternatives to Bleach
| Tool | Used in agent projects | Downloads / wk | GitHub stars | License |
|---|---|---|---|---|
| Beautiful Soup pypi Python HTML/XML parsing and screen-scraping library for web data extraction. |
160 (1.3%) | 69M | — | MIT License |
| Puppeteer npm Headless Chrome/Chromium automation for scraping, testing and screenshot generation. |
115 (0.9%) | 11.4M | 95.7K | Apache-2.0 |
| lxml pypi High-performance XML and HTML processing library combining libxml2/libxslt with Python. |
98 (0.8%) | 88.4M | 3,065 | BSD-3-Clause |
| Cheerio npm jQuery-like syntax for parsing and manipulating HTML/XML with server-side speed. |
94 (0.7%) | 26.6M | 30.5K | MIT |
| Playwright (Python) pypi High-level Python API for automating browser testing and web scraping. |
68 (0.5%) | 26M | 15K | — |
| fast-xml-parser npm High-performance XML parser and builder without native dependencies. |
58 (0.5%) | 83.3M | 3,139 | MIT |
| csv-parse npm Parse CSV files using Node.js streaming Transform API for memory-efficient processing. |
51 (0.4%) | 20.6M | 4,280 | MIT |
| Playwright Core npm Headless browser automation library for cross-browser testing and web scraping. |
49 (0.4%) | 122M | 97.4K | Apache-2.0 |