Scrape every FinCEN advisory, alert, notice, fact sheet, financial trend analysis, and SAR Activity Review PDF. Returns structured red-flag indicators, applicable industries, and typology tags plus full PDF text -- built for AML compliance and sanctions-screening teams.
Scrape the FineYoga (Fanyin 梵音瑜伽) registered teacher directory — China's largest premium yoga chain. Returns teacher names, workplaces, yoga specialties, and registration tier for every certified coach in the public registry.
Scrape the full FLOR modular carpet tile catalog. Extracts per-tile and box pricing, fiber content, pile height, face weight, eco-certifications, colorway names, and full product specs. Supports full-catalog, collection, designer, eco-only, and BYO URL modes.
Look up Florida-licensed insurance agents, agencies, adjusters and customer representatives by name, license number, or NPN. Returns license status, lines of authority, carrier appointments, and continuing-education standing, sourced from the Florida Department of Financial Services.
Scrapes the rotating featured FSBO listings from the ForSaleByOwner.com homepage. Returns up to ~20 active listings per run with address, price, beds, baths, sqft, property type, listing image URL, seller name, and detail URL.
Scrapes Brazilian developer job boards from GitHub Issues (frontendbr/vagas, backend-br/vagas and more). Parses free-text bodies into structured records with location, stack, seniority and employment type.
Scrapes the full FTC Legal Library cases-and-proceedings index. Returns structured case records with matter number, docket number, case status, type of action, bureau tags, case summary, and links to legal documents (complaints, consent orders, decisions).
Scrape enterprise-grade B2B software reviews from Gartner Peer Insights — verified-reviewer ratings, pros/cons, use case, reviewer role, industry, company size, and vendor responses.
Scrapes the Ghost Explore newsletter directory (explore.ghost.org), listing Ghost-hosted publications with member counts, descriptions, newsletter URLs, social links, and category tags. Covers ranking pages (Top Members, Top Revenue, Trending, Recent) and 30+ topic categories.
Scrape GSA eLibrary — all Schedule contract holders across every Special Item Number (SIN). Captures name, contract number, SIN, contact details, address, SAM UEI, socioeconomic indicators (SDVOSB/WOSB/HUBZone), and contract dates.
Enumerates the active business-license roll of HdL Companies municipal portals (Pomona, Hayward, El Cajon, Tustin and other reachable HdL cities) by business-type code, returning account number, business name, license status, expiry, and address for every license on file.
Search Sweden's hitta.se public person directory by name and get age, phone, and street address, plus municipality, county, and GPS coordinates, for private individuals. Built for debt-collection, tenant screening, insurance investigation, and B2C lead-list building.
Convert any URL or raw HTML to clean markdown, plaintext, and reader-mode metadata. Choose Readability (articles), Turndown (verbatim), or Trafilatura (noisy pages). Outputs title, byline, published date, language, word count, and reading time. FREE — the one-call URL-to-markdown primitive.
Extract live salvage-vehicle auction listings from IAAI's public search. Search by keyword or filter by state. Returns stock number, partial VIN, year/make/model, damage type, odometer, title type, run-and-drive status, sale date, branch location, and more.
Scrapes deforestation and fire data from INPE TerraBrasilis. Collects PRODES deforestation totals, DETER daily alerts, and fire records across Amazon, Cerrado, Pantanal, and other Brazilian biomes. Covers all states from 2015 to present. Essential for EUDR compliance and carbon offset research.
Scrape Instructables.com project metadata across all craft verticals including electronics, woodworking, sewing, cooking, and 3D printing. Returns title, author, category, difficulty, step count, materials summary, favorites, views, and comments counts per project.
Download and parse the IRS Tax Exempt Organization auto-revocation list. Returns all nonprofits that lost tax-exempt status for non-filing, including reinstatement records.
Scrapes the public JailbreakBench leaderboard tracking attack-success-rate for jailbreak techniques (PAIR, GCG, AIM, and more) against open- and closed-source LLMs, with and without defenses (SmoothLLM, perplexity filter, etc). Snapshot each run to track technique-vs-model ASR movement over time.
Job postings across Malaysia, Singapore, the Philippines, Indonesia, Hong Kong and Thailand — sourced from JobStreet and JobsDB (SEEK Asia). Get titles, companies, salaries, locations, classifications and full descriptions in one dataset.
Extract Kavak's owned used-car inventory across Mexico, Brazil, Argentina and Chile: price, financing, mileage, spec, hub location and availability for every listed vehicle, normalised into one cross-country schema.
Scrape live Kick.com streamer leaderboards by category — viewer counts, stream titles, and channel profiles (bio, follower count, social links). Pick specific categories or let it auto-discover the top trending ones each run.
Scrape Kickstarter campaigns across the whole site — all 15 categories and 154 subcategories — with funding goal, amount pledged, percent funded, backer count, creator history, launch date and deadline. Pick any mix of categories, or crawl everything.
Scrape classified listings from Kleinanzeigen.de (formerly eBay Kleinanzeigen). Extract listing details including title, price, description, location, seller info, images, and category-specific attributes.
Bulk-extract LeadIQ company profiles: industry, SIC/NAICS codes, employee bands, tech stack, key executives, continents of operation, and the observed corporate email-format pattern with confidence percentages. Covers roughly 1,045,000 companies with resumable crawls.