Generate realistic synthetic datasets with correlated fields, built-in presets (user profiles, companies, e-commerce products, log events), custom schemas, deterministic seeding, and multiple output formats (JSON, CSV, NDJSON).
Extract property listings from Property24 ZA and KE: price, bedrooms, agent name, EAAB registration, images, and 20+ fields per listing. Covers all 9 South African provinces and Kenya nationwide.
Scrapes the Bandcamp Discover feed for new releases. Filter by genre tag and sort order. Returns title, artist, label, genre tags, release date, price, and Bandcamp URL. Ideal for music supervisors, sync agents, and A&R researchers tracking indie new releases.
Scrape Canadian lawyer and law firm profiles from canadianlawlist.com. Extract names, firms, addresses, phones, emails, websites, and call-to-bar dates. Filter by province. Covers ~130k lawyers across all Canadian provinces.
Pull the CFTC Commitments of Traders (COT) report as structured rows. Covers legacy, disaggregated, and financial-futures variants. Filter by commodity, market code, and date range. Long, short, spread positions per trader category plus open interest and trader counts.
Scrapes the Grand Comics Database public API for comic issue bibliographic data — publisher, series, issue numbers, creator credits, story details, cover prices, and variant linkage. GCD is the canonical open comics catalog (2M+ issues). Data is CC-licensed.
Pull EU electricity-market data from the ENTSO-E Transparency Platform: day-ahead prices, load, generation by fuel type (nuclear, gas, wind, solar, hydro), cross-border flows, outages. All 61 EU bidding zones. No API key or signup — pick a dataset, zones, and date range.
Scrape attorney profiles from LawyerLegion.com. Extract names, contact details, practice areas, bar admissions, board certifications, education, and firm information. Filter by U.S. state. Ideal for legal marketing, recruiting, and lead generation.
Extract doctor profiles from Lybrate.com — specialties, qualifications, clinic info, ratings, consultation fees, and online availability across 200+ Indian cities
Scrape public investor profile metadata from PitchBook without a subscription. Supports text search, direct profile URLs, and bulk sitemap discovery. Returns firm name, description, location, investor type, status, investment metrics, social links, and more.
Scrapes all active freelance writing job listings from the ProBlogger Job Board — the canonical curated source for paid blogging, copywriting, editing, and ghostwriting gigs. Returns job title, company, type, location, pay rate, description, category tags, and listing URL.
Search and extract job listings from SEEK Australia by keyword, location, salary range, work type, and category. Returns job titles, companies, salaries, locations, work arrangements, and posting dates. Download as JSON or CSV — no login required.
Extract and validate structured data from any URL: JSON-LD, Open Graph, Twitter Cards, microdata, RDFa, meta tags. Local schema.org validation. Flags Google rich-result eligibility and AI-discovery readiness. Pure HTTP. Built for SEO audits and structured-data debugging at scale.
Scrape top-performing TikTok ads from the Creative Center. Filter by country, time period, industry, campaign objective, and sort order. Extracts ad creative metadata including video URLs, thumbnails, engagement metrics (likes, CTR, cost rank), brand names, and keywords.
Scrape towing and roadside assistance company profiles from Towbook's public directory (public.towbook.com). Extracts company name, address, city, state, zip, phone, and fax for every listed business.
Scrape architectural project data from ArchDaily — the world's most visited architecture website. Extracts project metadata, location, team credits, images, drawings, and publication details for thousands of global projects across all building types.
Scrape BoardGameGeek game data including rankings, ratings, categories, mechanics, designers, and more. Supports hot list, keyword search, and direct game ID lookup. Uses residential proxies to bypass BGG's datacenter IP blocks.
Scrape Bold.org, one of the largest scholarship marketplaces, into a structured database. Every record carries scholarship name, award amount, application deadline, education level, field of study, eligibility summary, category, sponsor organisation, applicant count and a no-essay flag.
Scrapes yoga, meditation, and ayahuasca retreat listings from BookRetreats.com for LATAM countries. Covers Mexico, Brazil, Peru, Costa Rica, Guatemala, Colombia, Nicaragua, Ecuador, Argentina, and Chile. Returns retreat name, center, location, price, rating, room options, and full description.
Search São Paulo Junta Comercial (JUCESP) by company name, CNPJ, or NIRE. Returns NIRE, CNPJ, company type, constitution date, share capital, business object, and address. Every term returns a row — a confirmed "not registered in SP" counts as a result. For KYC, AML, and M&A diligence.
National cannabis license database. Federates state cannabis boards into one normalized dataset: license number, type, status, business name, address, expiration. For cannabis B2B sales: software, payments, packaging, labs, insurance, M&A intel.
Scrape software listings and reviews from Capterra. Extract product names, vendors, categories, multi-dimensional ratings, pricing, deployment options, features, integrations, alternatives, and descriptions. Browse by software category or scrape specific product URLs.
Scrape job listings from Catho.com.br — Brazil's largest paid job board. Extract job title, company, salary, location, employment type, benefits, and full description. Ideal for salary benchmarking, labor market analysis, and job aggregators.
Track new token listings and delistings across Binance, OKX, Bybit, KuCoin, Kraken, and Coinbase. Unified schema with exchange, token symbol, listing type (spot/futures/margin), trading pairs, announcement URL, and timestamps. Built for traders chasing listing alpha and crypto data teams.