Scrapes Anytime Fitness club locations from their sitemap — covers 3,000+ US and international clubs. Extracts club name, address, contact details, coordinates, hours, and amenities from each location page.
First ATC actor. Crawls official Appalachian Trail Conservancy trail updates (closures, bear warnings, detours, alerts) and state section pages. Authoritative free AT logistics data for thru-hiker planning — shelters, sections, closures.
Extract German job listings from Arbeitsagentur.de (Bundesagentur für Arbeit), Germany's federal employment agency. Search by keyword, location, or browse all open positions. Returns job title, employer, location, contract type, posted date, and job URL.
Scrape race and sports event listings from Asdeporte, Mexico's leading endurance-sport registration platform. Returns events with name, date, location, modality, distances, price, organizer, and registration status.
Scrapes the ASPCA's canonical toxic and non-toxic plants database — the gold-standard pet-safety reference. Returns tri-species toxicity data (dog/cat/horse), scientific names, families, toxic principles, clinical signs, and images for ~1,024 plants.
Scrapes California State Bar disciplinary actions — disbarments, suspensions, probations, censures. One row per action with attorney name, bar number, sanction type, effective date, violations, and profile link.
Scrape property auction listings from Auction.com. Extract address, price, auction date, status, beds/baths, square footage, and more. Filter by US state or price range. Ideal for real estate investors, lead generation, and market research.
Scrape the Audible audiobook catalog — narrators, series position, Whispersync flag, pricing, runtime, and ratings. 20+ fields per title. Seeded from Audible's official product sitemap; no search keyword required.
Scrape Audubon's "Plants for Birds" database by US ZIP code — returns native plants ranked for bird habitat value with bird species attracted, plant type, and wildlife resources provided. Unique zip → native-plant → bird-ecology join not available elsewhere as a structured data feed.
Scrape the Auth0 Marketplace. Extract integration names, vendors, categories, feature types, tags, and descriptions across social connections, MFA, SSO, and security integrations. Built for identity/security ISVs and CIAM competitive intelligence.
Detect high-intent B2B sales triggers — hiring surges, funding rounds, executive changes, and news momentum — from a list of company names. Produces a graded (A/B/C/D) priority list with rationale and source signals.
Scrape Babylist's product catalog: product details, pricing, ratings, multi-retailer links, editorial badges, and category data from the leading US baby registry platform. Supports sitemap-driven full crawl, category browsing, editorial best-of list extraction, and direct product URL lookup.
Scrapes B&H Photo's Pro Audio catalog: microphones, studio monitors, audio interfaces, mixers, headphones, signal processors, and podcast gear. Extracts SKU, price, stock status, ratings, specs, and features. Supports used/B-stock inventory.
Scrape stolen-bike records from the Bike Index public API — manufacturer, frame model, serial number, colors, stolen date, location, and photo URLs. Filter by city, manufacturer, date range, or stolenness status for insurance, bike-shop fraud screening, and marketplace integrity use cases.
Scrapes Biketo (美骑网) — China's largest cycling portal — for news, product reviews, and race coverage since 2008. Enumerates articles by sequential ID across three channels. Returns title, author, publish date, channel, body text, lead image, and engagement metrics.
Scrape Japan's official court auction database (BIT / 不動産競売物件情報サイト) — all 50 district courts. Extracts case numbers, addresses, property types, areas, bid prices, deadlines, and 3-point-set PDF links for below-market foreclosure listings.
Scrape book series reading order data from bookseriesinorder.com. Returns publication order, chronological order, series positions, and book metadata for thousands of authors and series.
Scrapes active vehicle auction listings from Bring a Trailer (bringatrailer.com). Returns all current auctions with title, year, current bid, country, end time, and direct listing URL.
Scrape China Animal Health and Epidemiology Centre (CAHEC) disease-surveillance bulletins. Extracts full bulletin text, publication date, section, issuer, and NLP-tagged disease, species, and region terms from seven content sections including epidemiology surveys, policy notices, and center news.
Scrape property listings from Caixa Econômica Federal's real estate auction portal. Filter by state, city, modality and minimum discount. Returns appraisal value, discount %, address and edital link.
Scrapes the official Riotur carnival agenda at carnaval.rio — blocos de rua, camarotes, bailes, ensaios técnicos, and sambodromo events. Returns structured event records including title, date, location, event type, and description in Portuguese.
Scrapes the CBAt official sanctioned road-race calendar for Brazil. Returns each CBAt-permitted event with date, race name, distances, city, state, and permit level (Bronze/Ouro/Prata). Brazil is LATAM's top running market; CBAt is the sole sanctioning federation for ~1500 events per year.
Scrapes sanctioned cycling events and state federations from the Confederação Brasileira de Ciclismo (CBC). Returns event details, dates, locations, disciplines, and federation contact information.
Scrapes the CB Insights Complete Unicorn List — the canonical reference of all private companies valued at $1 billion or more. Returns company name, valuation, date joined, country, city, industry, and notable investors for every unicorn on the list (1,200+ companies).