Scrape and aggregate app reviews and ratings from both the Apple App Store and Google Play into one unified dataset. Filter by star rating. Great for ASO, market research, and review monitoring.
Turn any URL (HTML page or PDF) into clean, token-efficient Markdown with metadata, ready for an AI agent or LLM. Auto-detects PDFs and renders JavaScript-heavy pages when needed.
Scrape public Bluesky data via the official AT Protocol API: an account's posts, profile, followers, and following. No login required. Clean structured output for social research and monitoring.
One row per domain: registrar, registration and expiry dates, domain age, nameservers and DNS host, MX records and mail provider, plus a full SPF, DKIM, DMARC and MTA-STS audit with a plain answer on whether the domain can be spoofed.
Verify a list of email addresses against DNS. Checks syntax, MX records, disposable and role accounts, detects the mail provider, and suggests a correction for mistyped domains like gmial.com. No SMTP probing, no bounces on your reputation.
Scrape every ad a competitor runs on Google Search, Display and YouTube, from Google's official Ads Transparency Center. Each creative comes with how many days it has been running, so proven winners are obvious. Run it again and it reports what launched, stopped and changed.
Search and scrape Hacker News stories and comments via the official API. Filter by keyword, type, points, comments, and date. Clean structured output for research, monitoring, and LLM datasets.
Turn any web page or site into clean, LLM-ready Markdown and structured JSON for RAG, agents, and fine-tuning. Strips nav/ads/boilerplate; returns main content + metadata.
Crawl a website and generate an llms.txt file (the emerging standard that helps LLMs and AI agents understand your site). Optionally build llms-full.txt with full page content. Output saved to the key-value store.
Turn any PDF URL into clean, LLM-ready Markdown and structured JSON. Extracts text + tables + document metadata for RAG, agents, and fine-tuning. No AGPL components.
Watch any Shopify or WooCommerce store and get a row for every price drop or increase, sell-out, restock, launch and removal since the last run. Schedule it and pipe changes anywhere. First run builds a free baseline; you only pay per product checked and per real change found.
Scrape Product Hunt launches via the official API: name, tagline, description, votes, comments, topics, makers, and links. Filter by topic, date, and ranking. Clean JSON for research and trend tracking.
Crawl a website and get a per-page SEO audit: title/meta/H1/canonical/robots/Open Graph checks, image alt text, word count, and a prioritized list of issues with a health score.
Scrape any Shopify store's full product catalog via the public JSON API: titles, prices, variants, SKUs, images, tags, and stock. Filter across many stores at once. No login, no anti-bot.
Scrape questions and answers from Stack Overflow and any Stack Exchange site via the official API. Filter by tag, keyword, and sort. Clean text output, perfect for LLM/RAG datasets and dev research.
Find out what any website is built and sold on. Detects the CMS, ecommerce platform, analytics, advertising pixels, payment providers, chat widget, marketing and reviews apps, consent manager, CDN, server and frontend framework, each with the evidence behind it.
Validate EU VAT numbers in bulk against VIES, the European Commission's official service. Returns whether each number is registered and active, the trader name and address where the country publishes them, and a consultation number you can keep as proof of the check.
Scrape any WooCommerce store's product catalog via the public Store API: names, prices, sale prices, SKUs, stock, categories, and images. Filter by price, category, stock, and sale across many stores. No login.
Scrape and aggregate open jobs from any company's Greenhouse, Lever, or Ashby board into one unified, structured dataset. Filter by keyword, niche, location, and remote. Public data, no login.
Free API: turn any URL into clean link-preview metadata including title, description, Open Graph and Twitter card, favicon, canonical, and language. Powers link unfurling and preview cards. No login.
Turn a list of company websites into a clean B2B lead list. Extracts emails, phone numbers in E.164, postal addresses, PO boxes, VAT and company registration numbers and social profiles, verifies every email against DNS, and returns one row per company. Pay per company, never per page.
Scrape Twitch channels via the official Helix API: profile, live stream status and viewers, recent videos, and top clips. Look up channels by name or search by keyword. Clean JSON output.