GitStar
GitHub TrendingTopicsLanguages
/

Scraper GitHub Repositories

Explore popular GitHub repositories tagged “scraper”.

Compare stars, forks, and programming language using the same GitStar view as GitHub Trending.

RepositoriesGitHub topic
Ranked by:Stars

Trending Repositories

firecrawl/firecrawl

The context API to search, scrape, and interact with the web at scale. 🔥

TypeScript168,8809,438
huginn/huginn

Create agents that monitor and act on your behalf. Your agents are standing by!

Ruby49,8114,290
NaiboWang/EasySpider

A visual no-code/code-free web crawler/spider易采集:一个可视化浏览器自动化测试/数据采集/网页爬虫软件,可以无代码图形化的设计和执行爬虫任务。别名:ServiceWrapper面向Web应用的智能化服务封装系统。

JavaScript44,3655,371
iawia002/lux

👾 Fast and simple video download library and CLI tool written in Go

Go31,6293,316
cheeriojs/cheerio

The fast, flexible, and elegant library for parsing and manipulating HTML and XML.

TypeScript30,4531,711
feder-cr/Jobs_Applier_AI_Agent_AIHawk

Open source AI job application bot in Python: auto apply to jobs, with a tailored resume and cover letter for each posting.

Python30,1964,625
gocolly/colly

Elegant Scraper and Crawler Framework for Golang

Go25,4301,856
apify/crawlee

Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.

TypeScript25,4201,629
Evil0ctal/Douyin_TikTok_Download_API

🚀「Douyin_TikTok_Download_API」是一个开箱即用的高性能异步抖音、快手、TikTok、Bilibili数据爬取工具,支持API调用,在线批量解析及下载。

Python19,4102,728
getmaxun/maxun

🔥 The open-source no-code platform for web scraping, crawling, search and AI data extraction • Turn websites into structured APIs in minutes 🔥

TypeScript17,1781,476
codelucas/newspaper

newspaper3k is a news, full-text, and article metadata extraction in Python 3. Advanced docs:

Python15,1402,119
pwxcoo/chinese-xinhua

:orange_book: 中华新华字典数据库。包括歇后语,成语,词语,汉字。

Python11,6492,680
guyueyingmu/avbook

AV 电影管理系统, avmoo , javbus , javlibrary 爬虫,线上 AV 影片图书馆,AV 磁力链接数据库,Japanese Adult Video Library,Adult Video Magnet Links - Japanese Adult Video Database

PHP10,0302,003
apify/crawlee-python

Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Parsel, BeautifulSoup, Playwright, and raw HTTP. Both headful and headless mode. With proxy rotation.

Python9,441792
alirezamika/autoscraper

A Smart, Automatic, Fast and Lightweight Web Scraper for Python

Python7,882813
BruceDone/awesome-crawler

A collection of awesome web crawler,spider in different languages

7,279755
go-rod/rod

A Chrome DevTools Protocol driver for web automation and scraping.

Go7,063479
mishushakov/llm-scraper

Turn any webpage into structured data using LLMs

TypeScript6,913455
yujiosaka/headless-chrome-crawler

Distributed crawler powered by Headless Chrome

JavaScript5,636404
JustAnotherArchivist/snscrape

A social networking service scraper in Python

Python5,440781
AhmadIbrahiim/Website-downloader

💡 Download the complete source code of any website (including all assets). [ Javascripts, Stylesheets, Images ] using Node.js

HTML5,1631,183
niespodd/browser-fingerprinting

Analysis of Bot Protection systems with available countermeasures 🚿. How to defeat anti-bot system 👻 and get around browser fingerprinting scripts 🕵️‍♂️ when scraping the web?

JavaScript5,125273
bjesus/pipet

Swiss-army tool for scraping and extracting data from online assets, made for hackers

Go4,769219
fent/node-ytdl-core

YouTube video downloader in javascript.

JavaScript4,730854
d60/twikit

Twitter API Scraper | Without an API key | Twitter Internal API | Free | Twitter scraper | Twitter Bot

Python4,622561
joeyism/linkedin_scraper

A library that scrapes Linkedin for user data

Python4,429972
myreader-io/myGPTReader

A community-driven way to read and chat with AI bots - powered by chatGPT.

Python4,419438
UltimaHoarder/UltimaScraper

Scrape all the media from an OnlyFans account - Updated regularly

Python4,273615
IonicaBizau/scrape-it

🔮 A Node.js scraper for humans.

JavaScript4,076216
JavScraper/Emby.Plugins.JavScraper

Emby/Jellyfin 的一个日本电影刮削器插件,可以从某些网站抓取影片信息。

C#3,796558
GitStar

See what the GitStar community is most excited about today.

Trending

GitHub Trending TodayGitHub Trending WeeklyGitHub Trending Monthly

Languages

Browse all languagesTrending PythonTrending JavaScript

Explore

Browse GitHub topicsAI repositoriesDeveloper tools
© 2026 GitStarGitHub Trending source