Chemr is a document viewer; fuzzy match incremental search.
Cross-platform persistent and distributed web crawler :crab:
A crawler for scraping posts from medium.com
Shadowsocks. 科学上网, 仅供学习。是免费的服务器,可能存在科学上网不稳定。
A SoFIFA webcrawler and Machine Learning prediction
Python3 DHT 磁力种子爬虫 种子解析 种子搜索 演示地址
优雅地玩知乎
(deprecated) :cat: koshort is a Python package for Korean internet spoken language crawling and processing... or maybe Korean domestic cat.
sciBASIC# is a kind of dialect language which is derive from the native VB.NET language, and written for the data scientist.
网络数据采集技术—Java网络爬虫 (书稿完整代码,涉及网络爬虫的各种技术和知识点)
meijuba.net,Python crawler,M3U8格式视频下载,桌面应用
Google资深工程师深度讲解Go语言 爬虫项目。
爬取京东商品所有评论,利用情感分析,判断商品是否值得买
A series of distributed components for Scrapy. Including RabbitMQ-based components, Kafka-based components, and RedisBloom-based components for Scrapy...
万能小说下载器
Free Facebook pages MetaData Scraping Library - Unlimited Calls
Tapestry - 基于 Agent Skill Bundle 的轻量级书签知识库
解析视频 网站/APP/H5 页面视频信息。支持抖音、腾讯视频、YouTube、Instagram 等40余个网站与APP
Automatic accessibility checker with website crawling + screenshots for easy use
Oh no, stop this. You can see my local IP address 😲! Use `foundation` attribute against CRC32 lookup table to reveal local IP address of a Chrome/Chr...
Official Supadata MCP Server - Adds powerful video & web scraping to Cursor, Claude and any other LLM clients.
TumblTwo, an Improved Fork of TumblOne, a Tumblr Downloader.
Just a simple web crawler which return crawled links as IObservable using reactive extension and async await.
基于Nodejs,superagent,cheerio的在线web爬虫项目,支持生成API
A search engine for Open Data
获取滚动新闻
Crawl a website and take screenshots
🌌 High productivity semi-automatic crawler generator 🛠️🧰
一个基于 Tampermonkey 插件平台开发的爬虫。主要目的是最大限度模拟用户环境,避免被反爬虫系统识破。
A simple, open-source, easy to use, and free download manager for malware samples.
A Content Discovery and Development Platform. Empowering Cybersecurity, AI, Marketing, and Finance professionals and researchers to discover, analyze,...
Screen scraping and web crawling framework
Get the lyrics for the song currently playing on Spotify
Downloads news articles from Google news and uses pre-trained NLP models to perform sentiment analysis
Iota is a web scraper that can find all of the images and links/suburls on a webpage
🎧 Get json type billboard hot 100 chart
Spider ported to Node.js
Repo Python
Copy of http://phpcrawl.cuab.de/ for using with composer
日常代码爬虫、gui小工具等
基于go-gin框架建立减少冗余动作项目,如:下载一些工具
ProxyCrawl Python library for scraping and crawling
Crawl all your citations from Google Scholar
simple crawler for Korean banks with Transactions
Unfx Proxy Parser - Nextgen proxy parser with deep links crawler. Follow to internal links, third-party links. Sorting results by countries.
Serve clean Markdown from your Next.js site to AI agents, crawlers, and LLMs. Humans get HTML, agents get clean Markdown of the same pages. Two-file i...
talospider - A simple,lightweight scraping micro-framework
[Updated] A simple python crawler for my tutorial blog at http://www.jianshu.com/p/8fb5bc33c78e
Crawl Instagram hashtags
Scrape data from Google.com, Bing.com, Baidu.com, Ask.com, Yahoo.com, Yandex.com