Scrapes complete historical fight statistics from ufcstats.com. Retrieves per-fighter striking, takedown, and submission data for every recorded UFC bout. Ideal for MMA analytics, model training, and historical performance research.
Look up Wayback Machine snapshots for any URL or list of URLs. Returns capture timeline, optional snapshot markdown, and live-vs-snapshot diff. Date range filtering, capture limit, bulk input. Built for OSINT, journalism, SEO link-rot recovery, and legal evidence.
Scrape EUR-Lex for EU legislation: regulations, directives, decisions, judgments, treaties. Returns CELEX IDs, ELI URIs, EuroVoc subjects, amendment tree, legal basis, national transpositions, and full-text URLs. Built for LegalTech, compliance teams, and AI training pipelines.
Wrapper around tcgcsv.com — the nightly TCGPlayer pricing mirror. Covers 89+ TCGs: Magic, Pokemon, Yu-Gi-Oh, Lorcana, One Piece, Flesh & Blood, Star Wars Unlimited, and more. Walk sets or fetch specific cards. Sub-type-aware prices (Normal/Foil/Holofoil/1st Edition) with market/mid/low/high.
Scrapes independent insurance agency profiles from InsuranceDirectory.com (Big I / IIABA network). Extracts agency names, contact details, carriers represented, coverage specializations, licensed states, team members, and service areas across 60,000+ US independent agencies.
Scrape attorney profiles from LawyerLegion.com. Extract names, contact details, practice areas, bar admissions, board certifications, education, and firm information. Filter by U.S. state. Ideal for legal marketing, recruiting, and lead generation.
Scrapes new music releases from MusicBrainz, Apple Music charts, and Metacritic. Returns album title, artist, release date, type, label, Metacritic score, and source links — ideal for weekly release tracking, playlist curation, and music industry research.
Enumerate all ~25K public NYT Cooking recipes from the official sitemap and extract structured recipe data (ingredients, instructions, nutrition, ratings) from schema.org Recipe JSON-LD.
Extract and validate structured data from any URL: JSON-LD, Open Graph, Twitter Cards, microdata, RDFa, meta tags. Local schema.org validation. Flags Google rich-result eligibility and AI-discovery readiness. Pure HTTP. Built for SEO audits and structured-data debugging at scale.
Scrape statutory UK probate notices from The Gazette (London, Edinburgh, Belfast). Output: decedent and executor name and address, filing solicitor, date of death, creditor-claim expiry date. Native claim-expiry filter for heir-finders, dormant-account hunters, and estate solicitors.
Scrape Bold.org, one of the largest scholarship marketplaces, into a structured database. Every record carries scholarship name, award amount, application deadline, education level, field of study, eligibility summary, category, sponsor organisation, applicant count and a no-essay flag.
Scrape California state bid events and procurement contracts from caleprocure.ca.gov. Extracts event IDs, names, departments, bid types, dates, status, category codes, and contact information.
Scrapes CNINFO (cninfo.com.cn) — China's official filings portal for Shanghai, Shenzhen, STAR Market, ChiNext and Beijing exchanges. Returns annual reports, prospectuses, related-party deals and restated financials. Filter by exchange, date range, stock code or keyword. PDF URL per record.
Scrapes CJEU case law from the official CURIA juris search — judgments, Advocate-General opinions, and orders with ECLI/CELEX identifiers, parties, subject matter, and full-text/PDF links.
Scrapes JustWatch's public GraphQL API for streaming availability. Returns one record per offer: provider, format (SD/HD/4K), price, and deeplink. Covers Netflix, Prime, Disney+, Max, Hulu, Apple TV+, and 200+ more across 80+ countries.
Scrape 36Kr (36氪), China's premier tech and startup news outlet. Extract articles, funding-round announcements, author metadata, section tags, and full body text across all verticals.
Scrape the latest AI, ML, and data science job listings from foorilla.com/hiring (formerly aijobs.net). Extracts salary range, seniority, years of experience, remote policy, skills, education, tasks and apply URL. Covers the freshest ~105 public postings.
Scrapes retreat center profiles and reviews from AyaAdvisors, the leading directory for ayahuasca retreats. Covers legal-jurisdiction centers across Peru, Costa Rica, Brazil, Colombia, and Ecuador with ratings, reviews, pricing, and contact data.
Crawl toxic chemical release data from the EPA TRI via the Envirofacts API. Extract facility details, chemical names, release quantities by media (air, water, land), coordinates, and carcinogen flags. Filter by state, chemical, year, and facility.
Scrape the EU Transparency Register: organisations, financial disclosure, clients, accredited lobbyists, and Commissioner meetings. Ideal for public-affairs intelligence, ESG screening, political-risk research, and compliance due diligence.
Scrape ImportYeti's country directory to discover top suppliers and US importers by geography. Get ranked entity lists, recent shipments, and aggregate trade stats for any country — from national level down to individual cities and states.
Scrape the official LoL Esports schedule and match results for all leagues including Worlds, MSI, LCS, LEC, LCK and regional splits. Fetches live data directly from Riot's official lolesports API — the same source the official lolesports.com site uses.
Scrape The Met's open-access collection API. Returns full metadata for 490k+ objects: title, artist, department, medium, dimensions, classification, culture, credit line, AAT tags, and CC0 image URLs. Modes: keyword search, department walk, incremental, and direct object-ID lookup.
Download a bulk list of websites using WordPress. Returns clean, deduplicated site URLs — a few hundred or well over a million in a single run. A WordPress site database for agency prospecting, CMS market research, security surveys, and lead-generation pipelines that need the domains first.