← Back to Home

Scrapers & Parsers

Explore the highest-rated open-source web scraping tools and data extraction scripts. Sorted by GitHub authority, core language, and update frequency.

Fast, local-first web content extraction for LLMs. Scrape, crawl, extract structured data — all

A Powerful web scraper powered by LLM OpenAI, Gemini & Ollama

🔥 The open-source no-code platform for web scraping, crawling, search and AI data extraction •

🕷️ An adaptive Web Scraping framework that handles everything from a single request to a

The API to search, scrape, and interact with the web at scale. 🔥

通用 Discourse 论坛内容保存工具 - 支持全球数百个 Discourse 站点,一键保存到

PHP Library for detecting CMS

Instagram Bot which when given a post url will spam mentions to increase the chances of winning. Win

Public Roadmap for SerpApi, LLC (https://serpapi.com)

A simple python library that allows for easy access of the SEC website so that someone can parse

A powerful MCP server extension providing web search and content extraction capabilities. Integrates

HaiKei is an anime streaming website that uses the consumet API

Go cascadia package command line CSS selector

Web Data Scraper - no-code internet scraping. Extract and export to CSV, Excel, JSON, Google Sheets,

A web scraper for TikTok

A tutorial for web scraping using Playwright headless browser

Fast, lightweight Firecrawl alternative in Rust. Web scraper, crawler & search API with MCP server

MetaData html scraper and parser for Node.js (supports Promises only)

A command line program to download Hentai videos and images from multiple websites

A simple browser/client-side web scraper.

High-performance web crawler API optimized for LLMs. Turn any search or website into clean Markdown

Scrapes facebook's pages front end with no limitations & provides a feature to turn data into

A Reddit bot that summarizes news articles written in Spanish or English. It uses a custom built

A collection of awesome web scaper, crawler.