Hyperscrape Store
Hundreds of ready-to-run tools for web scraping and automation. Every one is real, maintained code — configure the inputs and go.
AI & LLM
Website Content Crawler
hyperscrape/website-content-crawler
Crawl an entire website and extract clean text for RAG & LLMs.
Wikipedia Article Scraper
hyperscrape/wikipedia-scraper
Structured Wikipedia data: summary, infobox, sections, links, images.
arXiv Paper Scraper
hyperscrape/arxiv-scraper
Search arXiv preprints and get authors, abstracts, categories and PDF links.
Crossref DOI Metadata Scraper
hyperscrape/crossref-scraper
Search 150M+ Crossref records: DOI metadata, citation counts, references and licenses.
OpenAlex Works Scraper
hyperscrape/openalex-scraper
Search 250M+ scholarly works on OpenAlex: citations, concepts, open-access status.
PubMed Article Scraper
hyperscrape/pubmed-scraper
Search PubMed biomedical literature and get rich article metadata as JSON.
Semantic Scholar Paper Scraper
hyperscrape/semantic-scholar-scraper
Search Semantic Scholar and get citations, open-access PDFs and rich paper metadata.
URL to Markdown
hyperscrape/url-to-markdown-scraper
Convert any web page into clean, faithful CommonMark Markdown with YAML front-matter.
Docs Crawler
hyperscrape/docs-crawler-scraper
Crawl a documentation section into per-page Markdown chunks, ready for RAG.
llms.txt Generator
hyperscrape/llms-txt-generator
Crawl a site's key pages and generate a ready-to-publish llms.txt file.
Readability Batch
hyperscrape/readability-batch-scraper
Batch-convert URL lists into minimal, uniform, LLM-ready text records.