Skip to content

好用的 Python 解析 HTML 库。写爬虫的小伙伴都感受过解析 HTML 的痛苦,常用工具 BeautifulSoup、lxml、Scrapy 的 selector 等。今天你有了新的选择 requests-html,支持 XPath、CSS 选择器、动态页面、过滤指定内容等。上手特别简单和迅速,我的爬虫项目 Hydra 中就用了它,解析 HTML 变得轻松了许多。Pythonic HTML Parsing for Humans™

active 2023-08-15 → 2026-09-13 (UTC)

Complete coverage27,369 / 27,369 hourly files (100%) · 2 absent upstream2023-08-15 → 2026-09-28 (UTC)
Events
1K
Pushes
0
Pull requests
23
Issues
42
Stars
807
Forks
71

Activity over time

Daily event counts in the loaded window

Line chart, 1126 days from 2023-08-15 to 2026-09-13. Pushes: 0 total, peak 0 in a day. Pull requests: 23 total, peak 5 in a day. Issues: 42 total, peak 2 in a day. Comments: 86 total, peak 4 in a day. Stars: 807 total, peak 5 in a day.

  • Pushes
  • Pull requests
  • Issues
  • Comments
  • Stars

Stars, PRs, issues and forks are under-captured in the later part of this window. GH Archive progressively stopped capturing non-push events during 2026 — −95% or worse by the end of the window. Every series here except Pushes fades for that reason, so a decline above reflects the archive, not this repository. Pushes stay reliable throughout, so read them, and the contributor counts derived from them, as the real signal. Data health has the measurements.

Top contributors

Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity

Recent activity

Latest issues, pull requests and releases

Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 807 stars here means stars gained during the window, not the repo's star count.