Scrape the QueryTracker literary agent directory — the #1 querying-author platform. Extracts agent name, agency, genres, query status, query method, social links, and last-updated date from every public agent profile. Ideal for query-CRM tools and literary data pipelines.
Query the official regulations.gov v4 API for federal rules, proposed rules, notices, dockets, and public comments. Filter by agency, document type, docket, date range, and keywords. Requires a free API key.
Crawl SEC EDGAR company data and filings. Extract tickers, CIK, SIC codes, addresses, and filing records. Search by ticker, company name, CIK, or SIC code. 800K+ companies, 12M+ filings.
Scrape ranked startup lists from SeedTable — startups across 4,000+ city lists worldwide. Returns startup name, industry, location, profile URL, and logo for every ranked company, and filters to just the cities you care about.
Scrape public profile metadata from Snapchat user profiles. Provide usernames or profile URLs and get display name, subscriber count, bio, avatar URL, and external links — one record per username.
Generate realistic e-commerce test data with interconnected products, customers, orders, and reviews. Features referential integrity, realistic distributions, temporal coherence, industry presets, and deterministic seed mode.
Fetch leagues, teams, players, and match events from TheSportsDB — the free multi-sport metadata API covering soccer, basketball, baseball, American football, and dozens more sports worldwide. Bring your own free or Patreon API key. Results are schema-normalised and dataset-ready.
Scrape business profiles and reviews from Trustpilot. Extract trust scores, ratings, star distributions, review text, and company details. Supports search queries, direct URLs, and category browsing.
Scrape software reviews and product data from TrustRadius. Extract ratings, trScores, reviewer details, pros/cons, and company metadata for competitive intelligence and market research.
Scrape full restaurant menus from UberEats. Extract restaurant info, all menu sections, items with prices, descriptions, and images.
Scrape beer check-ins (reviews), ratings, and metadata from Untappd. Input beer URLs or search queries to collect structured check-in data: ratings, reviewer info, comment text, serving style, venue, ABV, IBU, and beer style.
Scrapes the official US State Department travel advisory feed. Returns all 213+ country advisories with level (1-4), ISO code, risk indicators (Crime, Terrorism, etc.), summary, and dates.
Aggregates active ABC/liquor license data from US state portals — California (CSV), Texas and New York (Socrata API). Returns license holder, trade name, license type, status, address, and county. Ideal for distributor prospecting, compliance research, and POS system outreach.
Extract USAspending federal subaward records with full prime-contractor mapping: subawardee, prime contractor (name, UEI), awarding agency, amounts, NAICS, PSC, and place of performance. Built for federal contracting analytics, competitive intel, and small-business teaming discovery.
Scrapes Wimbledon draws and match scores from wimbledon.com internal JSON feeds — the same endpoints the official site uses. Covers all five draws (MS/LS/MD/LD/XD) with per-set scores, seedings, and court assignments. Defaults to current year; no login required.
Extract company profiles from the Y Combinator startup directory. Covers 5,700+ funded startups across all YC batches. Returns name, website, one-liner, description, team size, location, industry, batch, stage, hiring status. Filter by batch, industry, region, or hiring status.
Scrapes business listings from international Yellow Pages directories: Sweden (Eniro.se), Norway (Gulesider.no), and Spain (PaginasAmarillas.es). Returns unified records with business name, address, phone, email, website, category, rating, and opening hours.
Scrapes the ACE public exercise library. Extracts professionally-authored exercises with step-by-step instructions, body parts, equipment, difficulty, images, and attribution. For fitness apps, RAG systems, and health pipelines that require provenance-verified exercise content.
Scrape the Adobe Commerce (Magento) Marketplace. Extract extension names, vendors, USD list prices, edition and Magento-version compatibility, ratings, reviews, and install metrics. Built for Magento agencies, extension ISVs doing competitive pricing, and ecommerce platform researchers.
Extracts child well-being indicator data from the Annie E. Casey Foundation KIDS COUNT Data Center. Covers hundreds of indicators across all 50 US states and national level, multiple years. Fields: indicator name, location, year, data format, and numeric value.
Scrapes ai-bot.cn — the definitive curated China AI tool directory — into a clean dataset of 1,000+ tools with names, categories, descriptions, logos, and external links. Covers domestic (China) and overseas AI tools with origin classification.
Scrapes AKC-sanctioned dog show events and competition results from the American Kennel Club event search. Returns event details including venue, club, dates, superintendent, premium list and entry information for conformation, agility, obedience, rally, and other event types.
Scrape US intercity bus schedules and fares: Megabus and FlixBus (Greyhound/FlixBus network) with per-carrier trip IDs, origin/destination cities, departure/arrival times, fares, seat availability, amenities, and booking links. Covers 734+ city pairs and 500+ routes across the continental US.
Amtrak's full US passenger rail network: every route (Acela, Northeast Regional, Silver Meteor, Coast Starlight and ~50 others), 1,000+ stations with codes, state, timezone and amenity flags (wheelchair, staffed, QuikTrak), and current service alerts.