FF14 国服官网自动签到脚本
Headless web crawler for bugbounty and penetration-testing/redteaming
Doogle is a search engine and web crawler which can search indexed websites and images
Raven is a powerful and customizable web crawler written in Go.
CLI tool for fetching URLs from Wayback Machine, Common Crawl, and VirusTotal.
Production-ready web scraping in a single function call. Built on Crawlee.
The intelligence layer for any web scraper. Pair with Scrapling, Playwright, or httpx to crawl smarter.
对抗cloudflare载入页反爬虫防护(已失效)
自動抓取 ptt 表特板圖片
Crawler4U, a general purpose focused crawler
Web/FileSystem Crawler Library
📱 A utility for downloading Android apps from the Google Play Store and Xiaomi App Store (the Chinese App Store).
A simple python3 script used to download a users's friend list from facebook.
Papercut is a scraping/crawling library for Node.js built on top of JSDOM. It provides basic selector features together with features like Page Cachin...
A collection of webnovel sources offering varying amounts of scraping capability.
CobWeb is a Python library for web scraping. The library consists of two classes: Spider and Scraper.
Automatic proxy pool for web scraping - crawls, validates, and rotates proxies with rate limiting and MITM support
Get facebook events from location with Python 3
A multithreaded tool for downloading search results of Baidu image search.
Download de livros para PDF/EPUB - Integrada.minhabiblioteca / vitalsource
Generic altcoin DNS seeder. Compatible with virtually any cryptocurrency cloned from bitcoin. Built-in lightweight DNS server ~ Cloudflare DNS support...
🔎 scan the internet to find "private" proxies.
Best manga-viewer on windows for crawling/downloading/browsing exhentai.
將一個 E-Hentai 畫廊下載並轉換成 PDF,方便在 Kindle 上閱讀 以及在 iPad 上閱讀並作筆記,,,
✨ Article Crawler is a package used to crawl articles with Markdown format from a specific webpage and store them locally in HTML / Markdown formats.
A client implementation of Firefox DevTools over remote debug protocol in python
A collection of pentesting web scanners
基于小红书web端的请求封装,JS实现
Web Scraper and Crawler for LLM Apps and AI Workflows with NoCode / LowCode. Plug and play with your own logic and customize it flexibly and scalably...
[DEPRECATED] Simple, flexible, delightful web crawler/spider package
A Tox DHT network crawler
an unofficial facebook api
A collection of short projects, you could try and implement these as short projects or use them as part of a larger project.
🍋 Python基础、Pygame游戏编程、Python算法与面试题、四种常用的Python Web框架、爬虫、数据可视化、机器学习。一共七个Python大方向!
一个将runoob.com转换为PDF的爬虫
基于 Xray-core、glider 的代理池工具
华数杯2024C题数据集收集过程
Instagram Data Scraper analyze profile
Sneakpeek is a framework that helps to quickly and conviniently develop scrapers. It’s the best choice for scrapers that have some specific complex sc...
Awesome list dedicated to digital and data preservation tools, sources, services and so on.
Crawl any website into a single searchable file. Query it forever, offline.
HttpClient + Jsoup + Queue
마루마루 다운로더 신규 프로젝트
The fast website crawler
[Obsolete] imooc web crawler in Node.js(使用 Node.js 编写的慕课网爬虫)
🔥 Golang basics and actual-combat (including: crawler, distributed-systems, data-analysis, redis, etcd, raft, crontab-task)
:beetle:简单轻便的Java爬虫框架,只要会一点简单的正则表达式和简单的css选择器就能轻松的采集数据。
Simple for use node html crawler (spider) of site web pages
Fetches PubMed article IDs (PMIDs) from email inbox, then crawls PubMed, Google Scholar and Sci-Hub for respective PDF files.
Based on Swoole,a PHP DHT crawler, which have insane productivity(依托于swoole的PHP版本的DHT爬虫,有着奇高的效率)