Scrape crowd-sourced law school admissions data from LSD.Law — median LSAT and GPA, 25th/75th percentiles, acceptance rates, US News rankings, and application counts for all ABA-accredited law schools.
Scrape used heavy equipment listings from MachineryTrader.com across every major equipment category (excavators, skid steers, dozers, cranes, loaders, and more). Get pricing, full specs, dealer contact info, location, and photos for every for-sale listing.
Cross-ecosystem feed of confirmed malicious/typosquatted packages across npm, PyPI, crates.io, Go, Maven, NuGet, Packagist and RubyGems, sourced from the OSSF malicious-packages dataset (also feeds OSV.dev). Category taxonomy plus typosquat-target linkage for CI gating. Passive read only.
Scrapes the Mararun platform — the dominant Chinese marathon management SaaS. Returns event details for mararun-hosted Chinese marathons: name, date, city, registration windows, participant cap, organizer, and CAA/AIMS certification.
Scrape the complete Maskota.com.mx product catalog — Mexico's leading online pet retailer. Extracts product names, prices, brands, variants, tags, and category data for all pets across the Shopify-hosted store.
Search Meta Ad Library ads by keyword, advertiser page, country and ad type. Returns spend and impressions ranges, reach estimates, per-country reach, payer byline, state-media and AI-media disclosure flags, and creative-variant collation counts the standard scrapers omit.
Scrape the full MTGGoldfish Magic: The Gathering card price index. Extracts paper, online (MTGO), and foil prices for cards across all sets with weekly price-change data.
Extract live and completed government surplus auctions from Municibid — vehicles, equipment, and surplus gear from small-town and county agencies. Includes bid pricing, realized sold prices on ended lots, seller and location details, and vehicle specs like VIN and mileage.
Scrapes mid-career job postings from Mynavi Tenshoku (tenshoku.mynavi.jp), Japan's top-3 転職 board. Returns structured records with salary ranges, work hours, holidays/vacation, and required skills — useful for JP labor-market analysis, comp benchmarking, and recruiter research.
Scrapes the Mystic Stamp Company US catalog (mysticstamp.com). Returns 60k+ US postage stamp listings keyed to Scott catalog numbers — the standard US philatelic reference. Each record includes Scott number, title, issue year, denomination, price, stock status, and image URL.
Search Naver Place — Korea's dominant local directory — by region and category for business records: address, phone, hours, category, coordinates, visitor/blog review counts, booking/order flags, menu items, amenities, photos. Restaurants, salons, clinics, lodging, general businesses, nationwide.
Scrapes the National Dance Education Organization's college dance program directory, job board, and events. Returns institution names, locations, contact info, degree offerings, and faculty counts.
Scrape product listings from Newegg category and search pages. Extracts product title, brand, current price, was price, shipping, rating, review count, stock status, seller name (1P Newegg vs 3P marketplace), item number, and image URL.
Scrapes Noble Knight Games for out-of-print and collectible board games, RPGs, wargames, and miniatures. Extracts condition grades (New/Mint/NM/VG+/etc.), price per condition tier, publisher, category, stock, and condition price ladder. The go-to OOP/used tabletop catalog for collector valuation.
Scrape free US people data from Nuwber — full name, age, phone numbers, current and past addresses, and relatives. Search by first/last name with optional state filter. Bypasses Cloudflare protection via a real-browser (Camoufox) fetch.
Scrape business listings from PagesJaunes.fr (French Yellow Pages) by keyword and location. Extracts business name, address, phone, email, website, opening hours, rating, and geolocation. Supports category-based search across any city, department, or region in France.
Scrape Italian business listings from PagineGialle by category and comune — name, address, phone, category, rating, opening hours, and partita IVA/email/website/servizi from each listing's detail page.
Scrape listings from PayPay Flea Market (Yahoo! Flea Market), Japan's second-largest C2C resale app. Extracts item id, title, price, condition, seller info, and images from embedded __NEXT_DATA__ JSON.
Convert PDF documents into structured JSON. Extracts text, tables, and fields from any PDF URL, and can optionally run an AI structuring pass that turns raw text into clean, organized JSON ready for automation or analysis. No API key needed.
Scrape the official Penguin Random House publisher catalog for book metadata: title, author, ISBN, imprint, format, publication date, price, description, praise blurbs, and more. Search-seeded crawl into detail pages returns primary-source data not available from consumer aggregators.
PEP screening for AML/KYC. Streams 1.87M politically exposed persons from OpenSanctions (daily refresh), Wikidata, EU MEPs, US Congress, and UK Companies House PSC. FATF categories, family/RCA graph, three modes: ingest, fuzzy-match screening, new-PEPs diff.
Scrape unified podcast metadata from Podscan.fm — cross-platform podcast intelligence with contact emails, host names, Apple/Spotify IDs, episode counts, and social links. Covers 3M+ podcasts indexed across all major platforms.
Scrapes FIFA World Cup 2026 prediction markets from Polymarket — live probabilities, volume, liquidity, and outcome prices for tournament markets including outright winner, group results, top scorer, and player props.
Scrape press releases from PR Newswire — headline, date, author, and full text. Supports any category listing page as start URL. Ideal for journalists, PR professionals, and researchers who need to monitor corporate announcements at scale.