Yozh Scraper is a powerful, open-source web scraping toolkit built on Playwright & Taskiq + Redis for scalable data extraction. Effortlessly bypass anti-bot systems with real Chrome builds and Camoufox fingerprinting. Key Features: • Advanced Anti-Bot & Stealth Mode • Full-featured Web Crawler with SSE streaming • Ready-to-use parsing presets (Amazon, Google, etc.) • Native MCP Support for AI Agents & LLM workflows • Built-in Proxy Manager
Hey Product Hunt community!
I'm super excited to officially launch Yozh Scraper today! 🦔
We built Yozh Scraper out of frustration with the current data extraction landscape. Developers usually have to choose between outrageous paying subscription fees for commercial proxy APIs or spending endless hours building messy, custom infra that breaks the second an anti-bot system updates.
We wanted an open-source, high-performance toolkit that gives developers and data teams complete ownership over their web scraping pipelines—without the headache.
What sets Yozh Scraper apart:
Bypass Anti-Bot Shields: Powered by Playwright, real Chrome builds, and Camoufox fingerprinting to scrape modern web apps seamlessly.
Production-Scale Crawling: Uses Taskiq + Redis architecture for horizontal scaling, async execution, deduplication, and SSE streaming.
Native AI & MCP Support: Built-in Model Context Protocol (MCP) support so your LLMs and AI Agents can query and extract web data directly.
100% Open Source & Self-Hostable: Spin it up in minutes using Docker Compose—no vendor lock-in, no hidden costs.
You can check out our live data site atdata.cyberyozh.proor dive right into the codebase onGitHub.
We'd love for you to give it a spin! What anti-bot challenges or web scraping features are you dealing with right now? I'll be here all day to answer your questions and chat about features.
Happy scraping!