#scraper (30 Repositories)
Ranked open-source repositories tagged with #scraper, scored by pull request acceptance likelihood and maintainer engagement velocity.
31.2%
45.3h
30 repositories tagged #scraper
TeamNewPipe/NewPipeExtractor
NewPipe's core library for extracting data from streaming sites
maoserr/epublifier
Converts some webnovels to epub format
Solr159/JavBoss
开箱即用的本地 JAV/视频 刮削、管理、播放软件,支持命令行一键安装和 docker 部署。只需简单添加目录,即可打造你的私人 JAV/视频 媒体库,带给你顶级的浏览体验,懒人必备。| Your local JAV/video manager.
henrique-coder/perplexity-webui-scraper
An advanced, high-performance Python client, MCP server, and REST API for reverse-engineering Perplexity AI's WebUI.
slive777/OpenAver
Free open-source desktop app that scrapes JAV metadata and generates NFO + cover art for Jellyfin, Emby & Kodi. No Docker, no CLI — one-click install on Windows & macOS. 8 built-in sources + optional Metatube federation (30+ providers), actress collections, cross-language tag aliases, and a REST API for AI agents.
openzim/gutenberg
Scraper for downloading the entire ebooks repository of project Gutenberg
PxyUp/fitter
New way for collect information from the API's/Websites
kameleo-io/kameleo
Anti-detect browser for web scraping and automation. Engine-level fingerprint masking for Chromium and Firefox. Self-hosted, Docker-ready. Integrates with Selenium, Playwright, and Puppeteer via SDKs in Python, JavaScript, and C#.
sokomishalov/skraper
Kotlin/Java library and cli tool for scraping posts and media from various sources with neither authorization nor full page rendering (Facebook, Instagram, Twitter, Youtube, Tiktok, Telegram, Twitch, Reddit, 9GAG, Pinterest, Flickr, Tumblr, Coub, Vimeo, IFunny, VK, Odnoklassniki, Pikabu)
firecrawl/firecrawl
The context API to search, scrape, and interact with the web at scale. 🔥
ruippeixotog/scala-scraper
A Scala library for scraping content from HTML pages
Anakin-Inc/anakin
Open-source web scraping API. Turn any website into clean markdown or structured JSON. Anti-detect browser, proxy auto-selection, self-hosted. One command: make up
HDoujinDownloader/HDoujinDownloader
A general-purpose doujinshi and image gallery downloader
fredwu/crawler
A high performance web crawler / scraper in Elixir.
iawia002/lux
👾 Fast and simple video download library and CLI tool written in Go
rajhodedara/live-sport-plugin
A robust live sports scraping and streaming plugin designed for media centers. Aggregates real-time feeds and delivers seamless IPTV playback.
henson/Scraper
Tracking the most popular Github repos, updated daily.
danieldotnl/ha-multiscrape
Home Assistant custom component for scraping (html, xml or json) multiple values (from a single HTTP request) with a separate sensor/attribute for each value. Support for (login) form-submit functionality.
sunny9577/proxy-scraper
⭐️ A proxy scraper made using Protractor | Proxy list Updates every three hour 🔥
guyueyingmu/avbook
AV 电影管理系统, avmoo , javbus , javlibrary 爬虫,线上 AV 影片图书馆,AV 磁力链接数据库,Japanese Adult Video Library,Adult Video Magnet Links - Japanese Adult Video Database
go-rod/rod
A Chrome DevTools Protocol driver for web automation and scraping.
gigobyte/HLTV
The unofficial HLTV Node.js API
salimk/Rcrawler
An R web crawler and scraper
codelucas/newspaper
newspaper3k is a news, full-text, and article metadata extraction in Python 3. Advanced docs:
SlavyanDesu/BocchiBot
BocchiBot is a multipurpose WhatsApp bot using wa-automate-nodejs library!
yjl9903/AnimeGarden
動漫花園 镜像站 | 动画 BT 资源聚合站 | 动画 BT 资源开放接口
d60/twikit
Twitter API Scraper | Without an API key | Twitter Internal API | Free | Twitter scraper | Twitter Bot
cheeriojs/cheerio
The fast, flexible, and elegant library for parsing and manipulating HTML and XML.
crwlrsoft/crawler
Library for Rapid (Web) Crawler and Scraper Development