Extracts discussion threads from the Top-Law-Schools forums: admissions-cycle chatter, scholarship-negotiation threads, and candid school reviews across every board. Each record includes the thread title, board, author, post and reply dates, reply/view counts, and the first post's full text.
Scrape property listings from Trulia — for sale, for rent, and sold. Accepts search page URLs or a city/state location. Returns price, beds, baths, sqft, address, listing URL, and more.
Search and extract UK company data from the official Companies House registry. Returns company name, number, address, status, type, SIC codes, incorporation date, officers with roles and appointment dates, and previous names.
Scrapes all 226 UK FCDO (Foreign, Commonwealth & Development Office) country travel advisories from the GOV.UK Content API. Outputs structured records with alert status, safety, health, entry requirements, and other per-country sections. No auth required — pure public API.
Scrapes the complete UNESCO World Heritage List — all inscribed and tentative sites with geo-coordinates, cultural/natural category, inscription criteria, danger status, area, and states parties. Data sourced from the official UNESCO World Heritage Centre.
Search and scrape Upwork freelancer profiles by keyword. Extracts hourly rates, job success scores, total earnings, skills, location, Top Rated status, availability, and more. Filter by Top Rated, Top Rated Plus, or US-only freelancers. Supports pagination to retrieve large result sets.
Congressional financial disclosures for the House and Senate: Periodic Transaction Reports (stock trades) plus full annual disclosure schedules (assets, income, liabilities, positions, agreements, gifts, travel). Filing years back to 2008, amendments linked, STOCK Act filing-delay flagged per trade.
Monthly and annual TEU container throughput for major US port authorities, normalized into one schema.
Scrape outfit coordinates from WEAR (wear.jp), Japan's largest ZOZO-backed fashion community. Extracts tagged apparel items, brands, style tags, Wearista influencer status, engagement metrics, and body context. For JP fashion-trend and influencer-marketing datasets.
Capture full-page screenshots or PDFs from a bulk list of URLs. PNG, JPG, WebP, or PDF output. Desktop/mobile/tablet viewports, dark mode, wait-for-selector. Each URL gets a dataset row plus a binary in the run's key-value store. Built for CI snapshot tests, SEO audits, brand monitoring, and QA.
Remote job listings with salary data, from We Work Remotely. Returns job title, company, salary range and currency, category, job type, location restriction, tags and apply URL. Filter by category and job type. For remote job boards and recruiters.
Scrape job listings from any Workday-hosted careers page (*.myworkdayjobs.com).
Extracts job title, requisition ID, location, posting date, job description, and
apply URL from any company using the Workday recruiting platform — no login required.
Scrapes Australian business listings from Yellow Pages AU (Thryv/Sensis) by category and suburb/state — name, address, phone, ratings, and enriched detail (email, ABN, services, payment methods) on request.
Scrape YouTube videos, channels, and metadata without an API key. Search by keyword or scrape a full channel's video catalogue. Returns structured records with video ID, title, description, view count, like count, channel, duration, publish date, and thumbnail URLs.
Extracts the Netherlands' national care-quality directory: care organisations, named care professionals, and patient ratings with per-dimension subscores, free-text reviews, and provider replies. Includes BIG registration and AGB codes where published. Covers the full published sitemap corpus.
Scrape AllTrails trail data by keyword or location. Returns trail name, difficulty, distance,
elevation gain, GPS coordinates, rating, review count, photos, and full trail descriptions.
Bypasses DataDome bot protection using a real browser with residential proxy.
Scrapes the full bicycle product catalog from BikeShop.com.br — Brazil's dedicated bike e-commerce retailer. Returns MSRP, sale price, parcelamento (installments), stock status, brand, and category for all SKUs including Oggi, Nathor, and more.
Scrapes Cambridge Dictionary entries with full learner metadata: headword, CEFR level (A1–C2), UK and US IPA pronunciation, audio URLs, part of speech, guideword, definitions, and example sentences. Ideal for vocabulary apps, language-learning curricula, and NLP datasets.
Look up any CIRO-regulated Canadian financial advisor: registration status, approval categories, employment history, provinces licensed, and regulatory disciplinary history. The Canadian analog to FINRA BrokerCheck — search a bulk list of names or fetch a known report by ID.
Scrape handmade craft item listings from Creema, Japan's largest handmade marketplace. Extracts item title, creator name, category hierarchy, price, material, images, and favorite count from product detail pages across all major categories.
Scrapes EEOC enforcement press releases. Extracts employer name, settlement amounts, discrimination type, statute cited (ADA/Title VII/ADEA), district office, and case numbers — turning EEOC newsroom into structured employment-risk records.
Scrape Glama's MCP registry for 23K+ Model Context Protocol servers. Returns metadata, tool schemas, attributes (official, remote-capable), source repos, and SPDX license. Supports catalog, search, and single-server modes. Credits Glama per its API license; links each record to its listing.
Extract home-services pro business listings from HomeAdvisor's directory by trade and location: business name, rating, review count, years in business, screening status, and service area.
Scrape curated maker project metadata from Makezine.com — titles, authors, categories, tags, tools, materials, excerpts and thumbnails from 2600+ DIY projects.