Turn any website into a clean Vector Database pipeline. Bypasses bot protections to extract pure GitHub-Flavored Markdown. Features built-in AI Vision OCR for images, Delta Crawling to save costs, Custom Regex, and direct Pinecone upserts for RAG agents.
Scrape Reddit posts, comments, user profiles, and subreddits without logging in. Features an unblockable Hybrid Stealth Engine to bypass Cloudflare 403s. Built-in AI summarization, deep multi-sort crawling, media downloads, and native MCP support for AI Agents!
Bridge the web and your Obsidian vault! Scrapes articles, strips ads, downloads images locally, generates AI tags, and syncs perfectly formatted Markdown notes directly to your hard drive via the MCP protocol.