🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl!
-
Updated
Sep 4, 2026 - Python
🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl!
Agent for collecting, processing, aggregating, and writing metrics, logs, and other arbitrary data.
jsoup: the Java HTML parser, built for HTML editing, cleaning, scraping, and XSS safety.
新一代爬虫平台,以图形化方式定义爬虫流程,不写代码即可完成爬虫。
Light-weight, simple and fast XML parser for C++ with XPath support
Html Agility Pack (HAP) is a free and open-source HTML parser written in C# to read/write DOM and supports plain XPATH or XSLT. It is a .NET code library that allows you to parse "out of the web" HTML files.
Simple and fast HTML and XML parser
A fast & lightweight XML & HTML parser in Swift with XPath & CSS support
Command line tool to download and extract data from HTML/XML pages or JSON-APIs, using CSS, XPath 3.0, XQuery 3.0, JSONiq or pattern matching. It can also create new or transformed XML/HTML/JSON documents.
htmlquery is golang XPath package for HTML query.
豆瓣电影top250、斗鱼爬取json数据以及爬取美女图片、淘宝、有缘、CrawlSpider爬取红娘网相亲人的部分基本信息以及红娘网分布式爬取和存储redis、爬虫小demo、Selenium、爬取多点、django开发接口、爬取有缘网信息、模拟知乎登录、模拟github登录、模拟图虫网登录、爬取多点商城整站数据、爬取微信公众号历史文章、爬取微信群或者微信好友分享的文章、itchat监听指定微信公众号分享的文章
XPath package for golang, supports HTML, XML, JSON document query and more
To associate your repository with the xpath topic, visit your repo's landing page and select "manage topics."