Scrapes public Twitch data with no login or API key: channel profiles and follower counts, live stream status and viewers, top categories, clips, past broadcasts and search. Anonymous access is capped at 30 rows per query by Twitch, and every run reports it.
Extracts guest reviews for tours/activities listed on Viator.com -- currently in recon mode while DataDome's profile-rotation requirement is confirmed.
Product listings from Wayfair category and keyword pages: name, brand, price and was-price, rating, review count, stock message and promo flags. De-duplicated by SKU, with the resolved category reported.
Turns any website into a clean, embedding-ready corpus. Strips navigation, footers and cookie banners, converts the real content to markdown, and splits it into overlapping chunks that each carry their own URL, title and metadata. Crawl by URL list, sitemap or link graph.
Query the keyless Wikimedia Analytics API (AQS) for any Wikipedia project. Article and project pageviews over time, the top 1000 articles per day/month, views by country, unique devices, Commons media requests, and edit/editor activity. Legacy pagecounts back to 2007. CC0.
Wikipedia articles with daily pageview history and Wikidata facts attached. Flags the silent redirects Wikipedia never reports - asking for 'AI' returns 'Artificial intelligence' with nothing saying so - and reports the 10,000-result retrieval cap that sits under a far larger hit count.
Scrapes property listings from Willhaben, Austria's largest classifieds site, across rentals, apartments for sale, houses for sale and land. Every row carries price, price per m², rooms, living area, address, GPS coordinates, agency and images.
Scrapes any self-hosted WordPress site via its public REST API — posts, pages, media, categories, tags, comments, users, even custom post types. No login. Raw JSON passthrough, honest pagination, and a clear error row for sites that disable their REST API.
Macroeconomic and development data from the World Bank: 29,544 indicators across 217 countries. Flags the 78 aggregate rows the API mixes in with real countries, which otherwise overstate a global total roughly tenfold.
Scrapes YouTube search results, video details, channels, playlists and comments — views, likes, descriptions, durations, subscriber counts and full comment threads. No login or API key. Does not download video or transcripts (YouTube gates those behind a browser-only token).
Scrapes property listings for sale or rent from Zoopla, the UK's #2 property portal. Search by location, property type and price/bedroom filters; returns price, address, coordinates, features and agent from a single search call, with an optional detail pass.
Scrapes job listings with full descriptions from Bayt.com — the Middle East's largest job site. Keyword and city search across 25 countries (or all at once), plus direct job-URL lookup.
Scrapes houses, flats and land for sale or to rent from OnTheMarket — the UK's #3 property portal after Rightmove and Zoopla. Search any town, city or outcode with price and bedroom filters; returns price, address, bedrooms, EPC, floorplans, full descriptions, photos and agent details.
Scrapes apartments and houses for rent or sale from QuintoAndar, Brazil's largest rental platform. Search any city or neighbourhood; returns price, area, bedrooms, bathrooms, address, amenities and photos.
Products from Walmart search, category shelves and item pages. Treats PerimeterX's HTTP 200 block page as the block it is, keeps recommendation carousels out of your search results, and never reports Walmart's unstable match count as a total.
Scrapes the Y Combinator startup directory by industry. Each row carries company name, description, YC batch, location, logo and social links, with optional founders, team size, tags and open job postings.
Scrapes used-car listings from Encar, South Korea's largest car marketplace. Each row carries price, mileage, model year, manufacturer, trust/inspection badges and photos, with optional accident history, full options and VIN.
Scrapes events from Eventbrite by location, with optional keyword, date-range and language filters. Returns name, description, tags, venue with coordinates, times with timezone and ticket URL per event, plus an optional detail pass adding ticket prices, organizer and performers.
Scrapes property listings from Argenprop — one of Argentina's largest real estate portals (188K+ listings in apartments-for-sale alone). Returns price, currency, address, bedrooms, expenses, photos, agency and Argenprop's own internal location/type taxonomy ids for any search you paste in.
Scrapes flats and houses for sale or rent from Bien'ici, France's #3 real estate portal. Search any city, district or postal code; returns price, surface area, rooms, bedrooms, floor, energy class, full description, photos and seller details.
Scrapes job listings from jobs.adp.com, the careers site of ADP, one of the largest payroll/HR platforms in the US. Optionally filter by team; returns job title, url, posted date, location and the full description, address list and employment type.
Fetches full Bloomberg news articles from a list of URLs -- headline, byline, full body text, publish/update dates, topics, tags and lead image. No account or browser required.
Extracts guest review data (ratings, text, reviewer demographics, stay details, partner replies) for hotels listed on Booking.com by hotel URL or hotel ID.
Scrapes Flippa's public API for online businesses for sale: websites, e-commerce stores, SaaS, Amazon FBA, apps and domains. Returns asking price, monthly revenue and profit, traffic, verification flags, auction state and seller location, with honest reporting of Flippa's 10,000 total clamp.