Education scrapers
Education scrapers cover scholarly sources and open media catalogues. Openverse Media Scraper returns media rows, while OpenAlex Scholarly Works Scraper covers scholarly works.
172 scrapers, busiest first.
-
Openverse Media Scraper Scrapes openly licensed images and audio from Openverse by search term, media type, license, source, and aspect ratio. -
TMDB Movies, TV Shows and People Scraper Scrapes TMDB movie, TV show, and person records by search term or curated list (popular, trending, top-rated, now playin -
Hugging Face Papers Scraper Scrapes Hugging Face papers by search query or trending feed. -
OpenAlex Scholarly Works Scraper Scrapes OpenAlex scholarly works, authors, institutions, sources, concepts, publishers, and funders. -
DOAJ Open Access Journals Scraper Scrapes journal metadata from the DOAJ API. -
Crossref DOI Metadata Scraper Scrapes Crossref publication metadata by search query, title, author, or specific DOI. -
Niche Schools & Colleges Scraper Scrapes Niche ranking lists for K-12 schools, districts, and colleges. -
NSF Research Grants Scraper Scrapes NSF research grants by keyword, state, program, or date range. -
Udemy Scraper Scrapes Udemy courses, instructors, and reviews from course URLs or sitemap discovery. -
EU Clinical Trials EMA CTIS Scraper Scrapes EU clinical trials from the EMA CTIS register by search term, status, or phase. -
UN Comtrade Trade Statistics Scraper Scrapes UN Comtrade international trade records by reporter, partner, commodity, and year. -
Etymonline Word Etymology Scraper Scrapes word etymologies from Etymonline by keyword search or direct word lookup. -
ORCID Researcher Profile Scraper Scrapes ORCID researcher profiles by search query, institution, or ORCID ID. -
wger Exercise Database Scraper Scrapes exercises from the wger open fitness database. -
NIH RePORTER Scraper Scrapes NIH RePORTER research projects by keyword, agency, or fiscal year. -
College Scorecard Scraper Scrapes U.S. -
PyPI Packages Scraper Scrapes PyPI package metadata by exact name or keyword search. -
RePEc Economics Papers Scraper Scrapes economics papers from a RePEc series handle and returns each paper as a flat row with title, authors, abstract, -
Europe PMC Literature Scraper Scrapes Europe PMC biomedical literature by search query and returns each article as a flat row with abstracts, authors, -
Project Gutenberg Books Scraper Scrapes Project Gutenberg book metadata via the Gutendex API. -
TheCocktailDB Drink Recipes Scraper Scrapes cocktail recipes from TheCocktailDB by search or filter. -
Wikidata Entity Search Scraper Scrapes Wikidata entity search results by term, language, and type. -
GeoNames Places Scraper Scrapes GeoNames places, postal codes, and country dumps. -
Wikipedia Page Summaries Scraper Scrapes Wikipedia page summaries by direct title or keyword search. -
TheMealDB Recipes Scraper Scrapes recipes from TheMealDB by search, category, area, ingredient, or random. -
NASA Near-Earth Asteroids Scraper Scrapes near-Earth asteroid data from NASA's NeoWs API. -
WHO GHO Health Indicators Scraper Scrapes WHO Global Health Observatory indicators and returns each fact as a flat row with indicator code, country, year, -
OECD SDMX Statistics Scraper Scrape OECD economic and social statistics via the SDMX REST API. -
DEV.to Articles Scraper Scrapes DEV.to articles by tag, top week, latest, or username. -
IBGE Brazil Statistics Scraper Scrapes official Brazilian statistics from the IBGE SIDRA database by mode: localidades, aggregate catalogue, data obser -
Yu-Gi-Oh! Card Database Scraper Scrapes Yu-Gi-Oh! card records from the official database. -
Wikivoyage Travel Articles Scraper Scrapes full Wikivoyage travel articles by destination title or keyword search in any supported language. -
Docker Hub Container Images Scraper Scrapes Docker Hub container images by keyword search or direct owner/name lookup. -
Wikipedia Pageviews Scraper Scrapes Wikipedia pageview counts for a list of articles from the Wikimedia REST API. -
Open Library Editions Scraper Scrapes Open Library edition metadata by ISBN list or work key. -
GBIF Biodiversity Data Scraper Scrapes GBIF species search results and occurrence records. -
Zenodo Research Repository Scraper Scrape Zenodo research records by search query, community, or resource type. -
Open Trivia DB Quiz Questions Scraper Scrapes trivia questions from Open Trivia DB with optional category, difficulty, and type filters. -
Crates.io Rust Crates Scraper Scrapes Rust crate metadata from crates.io by search, name lookup, or sorted feeds. -
OEIS Integer Sequences Scraper Scrapes OEIS integer sequences by search query, A-number list, or keyword filter. -
Stack Exchange Q&A Scraper Scrapes Stack Exchange questions and answers by site, tag, search query, or date range. -
Metropolitan Museum of Art Scraper Scrapes artwork records from the Metropolitan Museum of Art's open-access collection by search query, department ID, or -
Figshare Research Data Scraper Scrapes Figshare research items by search, category, institution, or date. -
Dictionary Word Definitions Scraper Scrapes word definitions, synonyms, pronunciations, and etymologies from Dictionary.com and Thesaurus.com. -
REST Countries Reference Data Scraper Pull rich reference data on every country: official + common names, capital, currencies, languages, region, demonym, cal -
IETF Datatracker Documents Scraper Scrapes IETF Datatracker documents by type, working group, state, and date window. -
OSF Open Science Framework Scraper Scrapes public OSF research projects, preprints, and registrations by keyword, provider, or subject. -
HAL Open Science Scraper Scrapes research publications from HAL Open Science by query, document type, domain, or year. -
Europeana Cultural Heritage Scraper Scrapes Europeana's 60M+ cultural heritage items by search query and optional filters. -
Open Targets Platform Scraper Scrapes target, disease, and drug entities from the Open Targets Platform by ID or free-text search. -
Kaggle Datasets Scraper Scrapes Kaggle dataset listings by search term, tag, file type, license, or size. -
Stack Exchange Questions Scraper Scrapes Stack Exchange questions by site, search term, and tags. -
ONS UK Statistics Scraper Scrapes the ONS Beta API for UK statistics. -
INE Portugal Statistics Scraper Scrapes INE Portugal statistics by indicator code, theme, or catalogue. -
Statistics Canada Web Data Service Scraper Scrapes Statistics Canada's Web Data Service for catalogue entries, cube metadata, coordinate observations, and vector t -
UniProt Protein Sequence & Annotation Scraper Scrapes UniProt protein entries by search query or accession and returns each one as a flat row with gene names, organis -
OBIS Ocean Biodiversity Scraper Scrapes OBIS Ocean Biodiversity Information System data: occurrence records, species checklists, statistics, and dataset -
INE Spain Statistics Scraper Scrapes official INE Spain statistical operations, tables, datasets, and time series. -
RCSB PDB Protein Structure Scraper Scrapes RCSB PDB protein structure metadata by keyword search or PDB ID list. -
ChEMBL Molecules Scraper Scrapes ChEMBL molecules by name substring, ChEMBL ID, or molecule type. -
NCI GDC Cancer Genomics Scraper Scrapes NCI Genomic Data Commons public endpoints for projects, cases, files, or annotations. -
Argentina Open Data Scraper Scrapes Argentina's open data portal datos.gob.ar. -
EMA Medicines Scraper Scrapes the EMA medicines registry returning name, active substance, therapeutic area, status, marketing authorisation h -
USGS Water Services Scraper Scrapes USGS water data including instantaneous values, daily summaries, and groundwater levels by site code and paramet -
Tatoeba Sentence Corpus Scraper Scrapes Tatoeba sentences and translations by language pair, search term, or tag. -
REST Countries Info Scraper Scrapes country information from the REST Countries API. -
GBIF Occurrence Search Scraper Scrapes GBIF occurrence records by taxon, scientific name, country, year range, or record type and returns each record a -
Coursera Scraper Scrapes Coursera course listings by search query, returning ratings, review counts, difficulty levels, skills, and enrol -
ISS & Satellite Live Position Scraper Collects live position snapshots for the International Space Station and returns each one as a flat row with coordinates -
Drugs@FDA Approvals Scraper Scrapes drug approval records from the FDA's Drugs@FDA database. -
Internet Archive Search Scraper Scrapes Internet Archive items from a Lucene search query with optional filters for collection, media type, creator, and -
RFC Editor Index Scraper Scrapes RFC metadata from the RFC Editor index. -
KEGG Pathways Scraper Scrapes KEGG pathways, modules, and orthology entries by list, keyword search, or specific ID. -
Canada Open Data Catalog Scraper Scrapes Canada's open data catalog, returning each dataset as a flat row with title, description, organization, license, -
openFDA Medical Device Events Scraper Scrapes medical device adverse event reports from the openFDA database. -
openFDA Drug NDC Directory Scraper Scrapes the FDA National Drug Code Directory via openFDA. -
openFDA Drug Adverse Events Scraper Scrapes adverse event reports from the FDA's openFDA API. -
Colombia Open Data Scraper Scrapes any public dataset from datos.gov.co by its resource ID. -
MIT OpenCourseWare Scraper Scrapes MIT OpenCourseWare course listings by search query, department, or level. -
Wiktionary Definitions Scraper Scrapes Wiktionary definitions for a list of words from 10 language editions. -
Dutch CBS Statistics Scraper Scrapes Dutch CBS statistics tables and catalog entries. -
Nos Deputes France Parliament Scraper Scrapes French deputy profiles from nosdeputes.fr. -
openFDA Animal & Veterinary Scraper Scrapes animal and veterinary adverse event reports from the openFDA API. -
Open5e SRD 5.1 Scraper Scrapes Open5e SRD 5.1 entries by resource type, search query, challenge rating, spell level, or source document. -
Art Institute of Chicago Scraper Scrapes the Art Institute of Chicago's public collection. -
DailyMed FDA Drug Labels Scraper Scrapes FDA drug labels from the DailyMed database. -
Chile Open Data Scraper Scrapes dataset metadata and full rows from Chile's official open data portal, datos.gob.cl. -
DOAB Open Access Books Scraper Scrapes open access book records from the Directory of Open Access Books by search query, language, subject, or publishe -
Wikidata Lexemes Scraper Scrapes Wikidata lexemes by lemma search, language QID, or lexical category QID. -
PubChem Compound Scraper Look up chemical compounds on PubChem by CID, name, SMILES, or InChIKey. -
Australia Open Data (data.gov.au) Scraper Scrapes dataset metadata from Australia's official open data portal data.gov.au. -
D&D 5e SRD Content Scraper Scrapes spells, monsters, classes, equipment, and 20 other resource types from the D&D 5e SRD. -
LibriVox Audiobooks Scraper Scrapes the LibriVox free public-domain audiobook catalog. -
Open Targets Platform Scraper Scrapes targets, diseases, and drugs from the Open Targets Platform GraphQL API. -
Cat Breeds Scraper (TheCatAPI) Scrapes cat breed records from TheCatAPI. -
OpenAIRE Publications Scraper Scrapes open access research publications from OpenAIRE by search query and optional year range. -
Skillshare Course Scraper Scrapes Skillshare course listings by search keyword and returns each class as a flat row with title, instructor, studen -
edX Course Scraper Scrapes edX course listings by search keyword and optional subject filter. -
SWAPI Star Wars Scraper Collects Star Wars data from the open SWAPI by category and returns each record as a flat row. -
CORE Open Access Research Scraper Scrapes open access research paper metadata from CORE by keyword or title query. -
BioModels Models Scraper Scrapes BioModels model records by search query and returns each model as a flat row with ID, name, URL, format, submitt -
NCI GDC Cases Scraper Collects cancer case records from the NCI Genomic Data Commons API, filtered by project ID or primary site, and returns -
NASA Image and Video Library Scraper Scrapes the NASA Image and Video Library by search query. -
ChEMBL Compounds Scraper Scrapes ChEMBL compound records by clinical phase or molregno range. -
iNaturalist Observations Scraper Scrapes iNaturalist observations by taxon, place, and quality grade. -
OSF Projects Scraper Scrapes public OSF project listings by search query or research category and returns each project as a flat row with tit -
Ensembl Gene Lookup Scraper Looks up human gene symbols on Ensembl and returns stable IDs, biotype, chromosome location, and description as flat row -
NCBI Gene Database Scraper Scrapes NCBI Gene records by Entrez query and returns each gene as a flat row with identifier, symbol, name, location, a -
Figshare Research Articles Scraper Scrapes Figshare research articles by keyword search and item type filter. -
Zenodo Research Records Scraper Scrapes Zenodo research records by free-text search and resource type filter. -
WHO GHO Data Scraper Scrapes WHO Global Health Observatory data by indicator code. -
WHO GHO Indicators Scraper Scrapes WHO Global Health Observatory indicators by name filter. -
EBI OLS Ontologies List Scraper Scrapes the complete list of ontologies from the EMBL-EBI Ontology Lookup Service and returns each one as a flat row wit -
EBI Proteins API Scraper Collects protein metadata from the EBI Proteins API by protein name, organism, or both. -
UniProt Protein Scraper Scrapes UniProt protein entries from a search query and returns each one as a flat row with accession, protein name, gen -
bioRxiv Preprints Scraper Scrape bioRxiv and medRxiv preprint metadata by date range, subject category, or DOI. -
OpenStax Textbooks Scraper Scrapes OpenStax open textbooks by subject or search term. -
TCIA Collections Scraper Scrapes public cancer imaging collection metadata from The Cancer Imaging Archive. -
StatFin Finland Statistics PxWeb Scraper Scrapes statistical records from StatFin Finland's PxWeb database by subject path. -
SCB Sweden Statistics PxWeb Scraper Collects official Swedish statistics datasets from SCB's PxWeb API by subject path. -
bioRxiv and medRxiv Preprints Scraper Scrape preprints from bioRxiv and medRxiv by server and date range. -
SIMBAD Astronomical Objects Scraper Scrapes astronomical objects from SIMBAD using an optional ADQL where clause. -
NASA Exoplanet Archive Scraper Scrapes confirmed exoplanet data from the NASA Exoplanet Archive. -
Harvard Dataverse Datasets Scraper Scrapes Harvard Dataverse datasets, files, and dataverses by search query. -
EBI Ontology Lookup Service Scraper Scrapes ontology terms from the EBI Ontology Lookup Service and returns each term as a flat row with its IRI, label, ont -
UniProt Proteins Scraper Scrapes UniProt protein entries by search query and returns each protein as a flat row with accession, names, gene, orga -
OpenAIRE Publications Scraper Scrapes OpenAIRE publication records by keyword query and optional acceptance date. -
Europe PMC Scientific Literature Scraper Scrapes scientific publications from Europe PMC by search query. -
GBIF Species Taxonomy Scraper Scrapes GBIF species search by scientific name, keyword, or taxonomic rank. -
Europe PMC Articles Scraper Scrapes biomedical literature from Europe PMC by a free-text search query. -
NCBI Gene Summary Scraper Fetches NCBI Gene summary records for a provided list of Gene UIDs and returns each gene's name, description, synonyms, -
ChEMBL Assays Scraper Scrapes ChEMBL assay records by target ChEMBL ID, assay type, organism, keyword search, or direct assay ID. -
ChEMBL Targets Scraper Scrapes ChEMBL target records by search term, ChEMBL ID, organism, or target type. -
CTAN TeX Packages Scraper Scrapes CTAN TeX package metadata by keyword search. -
SMK Denmark Collection Scraper Scrapes artworks from the SMK Denmark collection by search term, artist, or public domain filter. -
Dryad Research Datasets Scraper Scrapes Dryad dataset metadata by free-text query or affiliation ROR ID. -
HPO Phenotypes Scraper Scrapes Human Phenotype Ontology terms by keyword or HPO ID, with optional disease and gene associations. -
Norway SSB Public Statistics Scraper Scrapes any public dataset from Statistics Norway's Statbank by table ID and returns each row as a flat record with offi -
VC Fellowships & Talent Networks Scraper Export fellows and talent network members from top VC programs (a16z, Neo Scholars, Pear Fellows, Kleiner Perkins, Gener -
Lithuania OSP Statistics Scraper Scrapes official statistical data from the Lithuania OSP public API by indicator code. -
Sigma-Aldrich Chemical Catalog Scraper Scrapes Sigma-Aldrich chemical products by search term, CAS number, or product number. -
Statistics Denmark Statbank Scraper Scrapes records from a specified Statbank table on statbank.dk and returns each row as a flat object with all dimensions -
Google Scholar Scraper Scrapes Google Scholar search results for a given query and returns each paper as a flat row with title, authors, public -
DrugBank Open Drug Reference Scraper Scrapes structured drug entries from the public DrugBank database. -
BCRP Peru Statistical Series Scraper Scrapes a single BCRP Peru statistical series by code and date range and returns each observation as a flat row with dat -
EU Open Data Portal Scraper Scrapes EU Open Data Portal datasets by search term. -
Art Institute of Chicago Artworks Scraper Scrapes Art Institute of Chicago artworks by search term. -
V&A Museum Artworks Scraper Scrapes V&A Museum artworks matching a search term and returns each one as a flat row with title, maker, materials, date -
Met Museum Artworks Scraper Searches The Metropolitan Museum of Art's public collection by keyword and returns each matching artwork as a flat row w -
Philippines PSA OpenStat Scraper Scrapes Philippine Statistics Authority OpenStat PXWeb catalog entries and table cell values. -
HKMA Daily Monetary Statistics Scraper Scrapes daily monetary statistics from the Hong Kong Monetary Authority for a specified date range. -
Statistics Canada Data Tables Scraper Scrapes Statistics Canada data catalogue by keyword and returns dataset metadata as flat rows. -
CBS Netherlands Statistics Scraper Scrapes CBS OpenData tables by table ID and returns each row of the TypedDataSet as a flat object. -
Opendata Swiss Datasets Scraper Scrapes Swiss open data datasets from opendata.swiss by keyword and returns each dataset as a flat row with metadata. -
disease.sh Global Health Data Scraper Scrapes global health data from disease.sh for every country. -
re3data Research Data Repositories Scraper Scrapes repository records from re3data.org by name or keyword and returns each one as a flat row with its metadata, sub -
Iceland Hagstofa PxWeb Statistics Scraper Scrapes data from Statistics Iceland's PxWeb API. -
Disease Ontology Terms Scraper Scrapes disease terms from the Disease Ontology by keyword or DOID. -
Niche Scholarships Scraper Scrapes Niche scholarship listings with award amounts, deadlines, eligibility text, and direct apply links. -
Wikimedia Commons Valued Images Scraper Scrape valued images from any Wikimedia Commons category via the MediaWiki API. -
Goodreads Book Details Scraper Scrapes Goodreads book details from lists, shelves, or search queries. - Smithsonian Volcanoes Scraper Scrapes volcano records from the Smithsonian Global Volcanism Program Holocene catalogue.
-
Preply Tutors Scraper Scrapes Preply tutor profiles by subject and optional country filter. -
Academic Research Aggregator Scraper Queries Crossref journals, OpenAlex institutions and topics, and Dryad datasets from a single keyword. - DOAJ Subject Classification Scraper
- Wikimedia Site Statistics Scraper
- PubMed Article Metadata Scraper
-
Crossref Academic Paper Metadata Scraper Scrapes academic paper metadata from Crossref by search query or publication type. -
Coloring.ws Printable Coloring Pages Scraper Scrapes coloring page images from Coloring.ws category pages and returns each image URL with its alt text, dimensions, a -
Wikimedia Commons Book Text Scraper Scrapes OCR text from DjVu and PDF books on Wikimedia Commons. -
Openclipart Line Art Scraper Scrapes Openclipart clipart by search query and returns each item with its title, artist, tags, download counts, SVG and -
Wikimedia Commons Subcategory Scraper Scrapes Wikimedia Commons subcategory names and IDs from a starting category.