scraper — independent software & tools
-
kari
Kari just a tui for getting media from providers and playing it in your media player with nice to have features.
-
social-media-scraper-skill
Extract and summarize social media content from platforms like Instagram, TikTok, X, and YouTube using Claude Code.
-
job-hunter
Automated job scraping pipeline for security engineer roles. Python + GitHub Actions + Gemini API.
-
ohara2
Self-hosted personal library for scraped web novels — read online or offline, no ads, no tracking.
-
ccrawl-cli
A fast, friendly command line for Common Crawl: URL index search, WARC fetch, Parquet columnar queries, and dataset building.
-
apify
The scalable web crawling and scraping library for JavaScript/Node.js. Enables development of data extraction and web automation jobs (not only) with headless Chrome and Puppeteer.
-
apify-client
Apify API client for JavaScript
-
@crawlee/core
The scalable web crawling and scraping library for JavaScript/Node.js. Enables development of data extraction and web automation jobs (not only) with headless Chrome and Puppeteer.
-
domparser-rs
A super fast html parser and manipulator written in rust.
-
firecrawl-cli
Command-line interface for Firecrawl. Scrape, crawl, and extract data from any website directly from your terminal.
-
google-news-url-decoder
A Node.js library to decode Google News URLs to their original source URLs.
-
google-play-scraper
scrapes app data from google play store
-
@mendable/firecrawl-js
JavaScript SDK for Firecrawl API
-
openbrand
Extract brand assets (logos, colors, backdrops) from any website URL
-
tiktok-live-connector
Node.js library to receive live stream chat events like comments and gifts from TikTok LIVE.
-
twitter-openapi-typescript
Implementation of Twitter internal API in TypeScript
-
url-metadata
Fetch a URL and scrape its metadata using Node.js or the browser. Can parse metadata from HTML strings or Response objects as well.
-
@vreden/youtube_scraper
A simple YouTube video downloader for audio and video formats with resolusi and quality.
-
Fentry-AutoCrate
This is a pre-configured script with GitHub Actions workflow to automatically claim the "Regular Crate" on fentry
-
Scrapling
Simplify web scraping by extracting data from modern websites with an easy-to-use Python library designed for efficiency and clarity.
-
Discord-Joiner-Token-Scraper
🤖 Automate Discord server joins using user tokens, streamlining bulk invitations while providing structured responses for each token processed.
-
sri-lanka-software-jobs
🚀 Automated Software Engineering Job Tracker for Sri Lanka. Scrapes and categorizes Intern, Associate, and SE roles daily using Java (Jsoup) and GitHub Actions. Centralizing the SL tech job market for developers.
-
API-Header-Spoofer
-
webustler
🌐 Extract clean, LLM-ready markdown from any URL, including Cloudflare-protected sites, with this effective MCP server for web scraping.
-
app-store-scraper
scrape data from the itunes app store
-
@bochilteam/scraper
Browserless scraper module
-
cheerio
The fast, flexible & elegant library for parsing and manipulating HTML and XML.
-
htmlmetaparser
A `htmlparser2` handler for parsing rich metadata from HTML. Includes HTML metadata, JSON-LD, RDFa, microdata, OEmbed, Twitter cards and AppLinks.
-
robots-parser
A specification compliant robots.txt parser with wildcard (*) matching support.
-
unfurl.js
Scraper for oEmbed, Twitter Cards and Open Graph metadata - fast and Promise-based
-
website-scraper
Download website to a local directory (including all css, images, js, etc.)