crawl pages to check what is for lunch today
头条号爬虫案例
Proxy pool. Finds and checks proxies with rest api for querying results. Can find over 25k proxies in under 5 minutes.
Fetch Iranian calendar events (Jalali, Hijri and Gregorian) from time.ir website
巴哈姆特自訂API
Web Scraping Framework
Serritor is an open source web crawler framework built upon Selenium and written in Java. It can be used to crawl dynamic web pages that require JavaS...
新浪博客文章/wenku8轻小说文库爬虫,可抓取图片保存,一键制作电子书。kindle读书党的神器!
🕸️ Spider Sitemap - Simple Python 3 crawler that automatically navigates your website, discovers all pages, and generates a complete XML sitemap. Easy...
Very simple bash script to crawl email addresses from a specific website.
This R package provides a crawler to scrape the European Energy Market EPEX SPOT at https://www.epexspot.com and the European Energy Exchange at https...
A distributed web crawler for xiaohongshu.com and visualization for the crawled content.
Better Google Dorking with Dorker.
An IP rotator via Tor for Scrapy.
对b站弹幕、评论进行爬虫,然后使用Word2Vec模型将其转化为词向量进行分析
Nodejs library that provides an Api for obtaining the movies information from FlixHQ website.
豆瓣影评爬虫助手 这个项目可以让你对感兴趣的电影进行影评数据抓取、分析。不仅可以看到影评的星级分布,还能查看根据点赞数加权后的平均星级,同时生成直观的...
Search Google Dorks like Chad. / Broken link hijacking tool.
A collection of WFDownloader scripts
instagram scraper tool automated insights
Telegram Bug Bounty Bot
Search Engine in Erlang
新浪微博爬虫:登录、关键词微博查询、微博监控
Google search results crawler, get google search results that you need - php
:trophy: Welcome to the wonderland of "AI" = f(DL, RL, DRL, ML, NLP, KG, MLOPS)
Competitive programming contests schedule
Tiny sitemap crawler for cache warming, and website status monitoring
This was the night of the crawling terror!
各种爬虫(目前支持Instagram、Weibo、Twitter)Miscellaneous crawlers (currently including instagram, twitter, weibo etc.).
A lightweight python wrapper designed for leveraging Google's search by image capabilities to perform reverse image searches programatically.
copymanga-downloader的mini ver,专为nas设计,不止于copymanga,支持多种平台!支持Web管理以及Docker部署!(当前支持copymanga、泰拉记事社、antbyw、gangano...
A web crawler based on requests-html, mainly targets for url validation test.
[deprecated] 유세인트 파이썬 클라이언트
Collect email addresses by crawling search engine results.
关于5000+站点的scrapy爬虫开发,涉及一些技术架构搭建以及各种反爬方案,详见readme文件
Powerful C++ web crawler based on libcurl
The latest way to get bet365 data odds, with a delay of 0.2 seconds bet365api
Crawler for fetching information of US Patents and PDF bulk download
little python projects, 一些小的python项目.
Recursive and multi-threaded broken link checker
抓取豆瓣小组相关信息(小组、用户、帖子)。
Hide your IP with free proxies using Froxy 🔄
This .NET Standard package provides convenient access to the Local API REST interface of the Kameleo Client.
来自[码云](https://gitee.com/generals-space/site-mirror-go) 通用爬虫, 仿站工具, 整站下载
Asynchronous file scanner and downloader for FTP servers.
GitHub Action to check a website for broken links
A utility to collect data from github stargazers, subscribers and contributors of a selected project
Download reddit posts based on keywords and perform sentiment analysis on the posts.
Web scraping script written in python using scrapy library in order to scrape product data from popular Sri Lankan vehicle selling web sites.
Unified search, crawling, and archiving toolbox for AI agents and automation scripts.