Scrape posts from any beehiiv-powered newsletter. Input publication domains — the actor discovers post URLs via sitemap and extracts title, author, publish date, excerpt, cover image, tags, and word count. Supports multi-newsletter fan-out in a single run.
Unified, machine-readable feed of public bug-bounty program scopes across HackerOne, Bugcrowd, Intigriti and YesWeHack, normalized into in-scope/out-of-scope asset arrays with run-over-run scope-change detection.
Scrape live and sold car auction listings from Collecting Cars, a global (UK, EU, US, AU) enthusiast vehicle auction platform. Get make, model, year, VIN, mileage, current bid or final sale price, reserve status, bid count, seller, highlights, and photos for every listing.
Company intelligence database from Craft.co: revenue, market cap, ESG and cybersecurity ratings, key executives, office locations, financials, operating metrics, competitors and subsidiaries per company.
Scrape all 550+ dog breed profiles from DogTime.com, including 25-axis trait scores (adaptability, friendliness, trainability, and more), vital stats, breed group, and detailed text sections for each breed.
Scrapes the SBTi (Science Based Targets initiative) public dataset. Returns 14,000+ companies with net-zero commitments, temperature alignment (1.5°C/2°C), target scope, ISIN, sector, and SBTi status. No login required.
Scrape Grailed designer and streetwear resale listings — active and SOLD — via the Algolia search API. Returns full listing data including sold price, seller score, original price, and arbitrage-ready comparables.
Extract Growjo's fastest-growing private companies with estimated revenue, headcount growth %, total funding, industry, HQ location, and CEO name — company growth data built for lead lists, VC sourcing, and market research.
Extract probate and civil court records from Harris County (TX) District Clerk. Search by party name or date range. Returns case info, parties, attorneys, and filing event history.
Scrapes haken (派遣) and temp staffing job listings from hatarako.net — Japan's leading dispatch job aggregator. Extracts job title, staffing agency, hourly wage, occupation, location, work hours, contract type, start date, required skills, and job URL across all prefectures.
Scrape Kenbiya (建美家 / kenbiya.com) — Japan's #2 investment-property portal
after Rakumachi. Extracts yield, monthly rent, occupancy, structure, and
broker data from ~20-40K active listings. Natural companion to
rakumachi-investment-scraper for full Japan investor-market coverage.
Scrape live listings from magi (magi.camp) — Japan's leading C2C trading-card marketplace. Search by keyword (Pokémon, Yu-Gi-Oh!, One Piece, MTG, etc.) and collect listing prices, sold signals, favorite counts, badges, and image URLs.
Extract European company records: registry identity, officers with appointment dates, multi-year published financials, merger and acquisition events, and register filings. Covers registers across Germany, Austria, Switzerland, UK, France, Benelux, and the Nordics.
Normalizes published prompt-injection and jailbreak datasets from HuggingFace and GitHub research repos into one labeled corpus: technique, target model, defense bypassed, license, cross-source dedup. Defensive only — aggregates public data for guardrail/eval testing, never targets a live LLM.
Extract government surplus auctions from PublicSurplus.com across every browse category - vehicles, heavy equipment, computers, furniture, industrial gear and more - from roughly 9,000 US and Canadian government sellers, with agency, location, bid, and sold-price data.
Scrape authenticated pre-owned luxury listings from RECLO, Japan's major luxury consignment marketplace. Returns product ID, title, brand, price in JPY, RECLO condition rank (N, S, A, B or C), product URL and primary image for every listing.
Scrapes the complete product catalog from REP Fitness and Titan Fitness — the two dominant direct-to-consumer strength-equipment retailers on Shopify. Returns normalized product records with titles, variants, pricing, availability, and images from both sources in a single dataset.
Scrape live job listings from Rozee.pk, Pakistan's dominant job board. Returns title, company, city, salary range, experience level, education requirement, functional area, industry, gender preference, skills, and full description for every posting in the unfiltered index.
Scrape curated highlight stories from public Snapchat profiles. Provide usernames and get direct media URLs, thumbnails, story titles, and creator metadata — one record per story snap.
Scrape job listings from StepStone across Germany, Austria, and Belgium. Get structured data on titles, companies, locations, salaries, work mode, contract type, and full job descriptions.
Scrapes highway and bridge construction bid-letting calendars from US state DOTs and federal SAM.gov. Covers upcoming bid openings, awarded contracts, engineer estimates, and DBE goals. Sources: Florida DOT and SAM.gov NAICS 237xxx solicitations.
Build an attorney contact list or law firm database from Justia.com. Extract names, contact details, practice areas, education, bar admissions, ratings, and reviews. Filter by U.S. state and practice area. For legal marketing, recruiting, and lead generation.
Scrape attorney and law firm data from FindLaw Lawyer Directory to generate high-quality, targeted legal industry leads
Build a local business database from Manta.com US small business listings. Extract names, addresses, phones, websites, emails, categories, revenue ranges, and employee counts.