Yozh Scraper icon
Yozh Scraper icon

Yozh Scraper

An open-source, Playwright-powered web scraping and crawling toolkit with native proxy integration, fingerprint spoofing, and LLM-friendly extraction.

Yozh Scraper screenshot 1

Cost / License

  • Free
  • Open Source (MIT)

Application type

Platforms

  • Mac
  • Windows
  • Linux
  • Self-Hosted
  • Docker
0likes
0comments
0articles

Features

Properties

  1.  AI-Powered

Features

  1.  Ad-free
  2.  Extensible by Plugins/Extensions

Yozh Scraper News & Activities

Highlights All activities

Recent activities

Yozh Scraper information

  • Developed by

    RS flagCyberYozh-data
  • Licensing

    Open Source (MIT) and Free product.
  • Written in

  • Alternatives

    7 alternatives listed
  • Supported Languages

    • English

AlternativeTo Category

Development

GitHub repository

  •  87 Stars
  •  11 Forks
  •  0 Open Issues
  •   Updated  
View on GitHub

Popular alternatives

View all
Yozh Scraper was added to AlternativeTo by Data Cyberyozh on and this page was last updated .
No comments or reviews, maybe you want to be first?

What is Yozh Scraper?

Yozh Scraper is a scalable, open-source web scraping and crawling toolkit developed by CyberYozh, designed for reliable data extraction from complex, dynamic, and anti-bot protected websites. Built on top of Playwright and Node.js/Python, it acts as a self-hostable bridge between automated headless browsers and dynamic web scraping workflows.

Unlike standard scraping libraries, Yozh Scraper comes pre-configured to bypass modern detection mechanisms out of the box. It integrates anti-fingerprinting technologies (such as Camoufox profiles and real Google Chrome engine support), automatic timezone/locale alignment, and rotating residential or mobile proxy integration.

Key Features: Headless & Real Browser Rendering: Seamlessly switch between Playwright Chromium, Camoufox, or real Google Chrome instances to mimic legitimate user behavior.

Advanced Evasion & Spoofing: Built-in hardware fingerprint randomization (GPU, OS, screen resolution, canvas) and native navigator getters.

Proxy & Geo-Alignment: Native compatibility with CyberYozh residential and mobile proxies, automatically matching browser timezones and locales with the exit IP location.

Full Website Crawler: Features a dedicated crawler engine (Yozh Crawler) that walks sites from seed URLs, deduplicates links, respects rate limits, and streams live results via Server-Sent Events (SSE).

Session & Authentication Support: Declarative login flows and cookie management enable scraping behind authenticated paywalls and social logins (e.g., LinkedIn, Amazon, eBay).

Pre-built Presets & AI Self-Healing: Ships with optimized recipes for major e-commerce and social sites, with support for LLM-based CSS/XPath self-healing.

MCP & AI Agent Integration: Includes Model Context Protocol (MCP) server endpoints, making it easy to plug Yozh Scraper directly into LLM frameworks and AI agents (such as LangChain, N8N, or custom agents).

Visual Tester UI & REST API: Comes with an easy-to-use Node.js web dashboard (scraper-tester) for testing single-page scrapes, batch jobs, and site mapping.