Crawler Configuration
v3.0.0Simulation simulates complex depths, outbound scopes, and OCR targets inside our sandbox.
Scans embedded images using Tesseract OCR to extract text/email signatures.
Downloads and parses attached PDF, DOCX, DOC, and TXT files found on pages.
Ignores web page text and images. Strictly crawls pages to locate & extract emails from PDF/Word document files only.
Live Dynamic Domain & OCR Engine Mapper
Engine idle... Ready to parse.Autonomous Server Daemons (24/7 Background Engine)
These processes run independently on your PHP / MySQL server. They continue scraping even when your browser tab is closed or your laptop is shut down.
Extracted Email Contacts Dataset
Aggregated email matches from HTML parsing and Tesseract OCR decodes. 0 Unique / 0 Total
| Match ID | Found Email Address | Discovery Source URL | Extraction Type | Context / Sector | Depth Level |
|---|---|---|---|---|---|
| No target crawls initiated yet. Configure and trigger the engine on the dashboard to fetch data. | |||||
Scraping History & Saved Session Archives
Full archive of all past scraping runs. Inspect collected emails, resume paused sessions, or export dataset backups.
Local Scraping Database Audit
Reports standard mock SQL jobs tracking, timestamps, success vectors, and OCR count metrics.
| Job ID | Target domain | Trigger Date/Time | Max Depth | Extracted Emails | Duration | Engine Outcome |
|---|