Apex Automation Team · Serving Clients Worldwide

Business Process Automation
Software QA Web Scraping

We run the whole process Client onboarding, portal and dashboard operations, document and OCR pipelines, verification, monitoring and reporting, all handled by bots and AI agents instead of your team.

And we test what you ship Two ways: a trained manual testing team that runs cycles on demand, or an automated testing system you own. Behind both sits the data layer: browser bots, custom web scrapers, bulk downloaders and verified lead databases.

Every solution is developed specifically for your target website and business requirements.

Client onboarding Portal & dashboards Document & OCR Verification & matching Monitoring & alerts Reporting & CRM sync
scraper_engine.js
$ node index.js --target=creator-marketplace
[info] Initializing Puppeteer cluster...
[info] Bypassing Cloudflare protection... Success
[info] Navigating to target directories...
▶ Scraped page 1/45 [Found 48 influencers]
▶ Scraped page 2/45 [Found 96 influencers]
[action] Downloading media assets: 14% [||.....]
[info] Extracting emails from social bio hooks...
✔ Data structured & validated. [342 items]
$ export output.xlsx done
Business Process Automation

We Automate Whole Business Processes

Not one-off scripts We build end-to-end workflow automation: one system that picks up the work, applies your business rules, routes approvals to the right person, handles exceptions instead of stalling, keeps every connected system in sync and records what happened. The same pattern fits client onboarding, order processing, invoicing and approvals, lead routing, ticket triage, and document or compliance checks. Your team sets the rules and signs off on what matters. The system runs the rest.

Operations OS

Client Engagement & Workflow Orchestration

Intake forms, task routing, status tracking, deliverable handoff and client reporting running as one orchestrated system, the same architecture as our Enterprise Operations OS sample project.

Portals & dashboards

Portal & Dashboard Operations

Bots that log into supplier, reseller, insurer or client portals to pull records, fill forms, place and update entries and refresh dashboards on a schedule, with logins, sessions and 2FA handled.

Documents

Document and OCR Pipelines

Bulk PDF processing, OCR, classification, splitting and renaming, then routing every file and extracted field into the right folder, sheet or system automatically.

Data quality

Verification & Data-Matching Workflows

Cross-checking records between systems such as VIN checks, invoice and order matching, duplicate detection and conditional formatting, with exception flags instead of silent errors.

Monitoring

Live Monitoring & Alerts

Watching sheets, sites, inboxes and channels around the clock and raising an email or Slack alert the moment something changes, breaks or leaks.

Reporting

Reporting and CRM Sync

Scheduled reports and two-way sync with Google Sheets, HubSpot, ClickUp, Airtable, n8n, Zapier and Make, or straight into your own database and API.

Solutions Spectrum

What We Can Build

The building blocks Workflow automation, browser bots, data extraction and AI agents behind those systems, engineered for speed, durability and correctness.

Web Scraping

  • Lead Extraction
  • Infinite Scroll Handling
  • Pagination Management
  • CSV / Spreadsheet Export
  • Excel Multi-sheet Formatting

Browser Automation

  • Auto Click Workflows
  • Form Input Automation
  • Dashboard Automation
  • Auto Login & Session Persistence
  • Report & Document Generation

Bulk Download

  • PDF & Document Downloader
  • High-Res Image Downloader
  • Video & Media Downloader
  • Automated ZIP Archives
  • Dynamic Asset Renaming

API Extraction

  • Hidden API Discovery
  • JSON Structured Export
  • AJAX Data Extraction
  • Network Call Analysis
  • Direct API Token Mining

AI Agents

  • Autonomous Workflows
  • n8n AI Agent Workflows
  • LLM API Integrations
  • Intelligent Chatbots
  • Automated Content Ops
  • Decision-making Loops

Custom Integration

All these are samples. We do work on every automation software you want.

  • Legacy Desktop & Web Systems
  • Custom RPA Workflows
  • Tailored Data Pipelines
Proven Demos

Featured Client Solutions (Selected Samples)

Real client work Selected samples, from complete business processes (operations orchestration, portal automation, PDF and OCR pipelines, verification workflows, live monitors and dashboards) to focused scrapers and lead pipelines. Any process or target website can be automated; our team reviews, tests and runs a free feasibility check to determine the exact scope.

Overview

Custom-built browser automation and API interception system that extracts detailed influencer metadata (handles, bio links, contact emails, stats) from Influencers Club. Can execute with a UI-based browser panel or via headless backend API parsing.

Key Features

  • Automated scrolls and session hooks
  • Dual modes: Graphical Web Scraping panel / Console Backend API data
  • Clean extraction of verified email lists
  • Structured table exports to Excel (XLSX)
  • Throttling engine to prevent account flags

Technologies

JavaScript Browser Console DOM Parsing API Analysis Chrome Ext

Result

Extracts structured creator information within minutes, bypassing typical rate-limits.

Select Run Mode:
Influencers_Club_scraped.xlsx
A B C D E
Username Followers Verified Email Country Niche Category
laura_fit 145,000 laura.fit@gmail.com USA Fitness & Wellness
cooking_with_sam 89,400 sam.colabs@cookingsam.co UK Food & Cooking
tech_guru_pro 320,000 info@techgurupro.net Canada Technology
beauty_by_ana 210,000 ana@beautybyana.com USA Beauty & Fashion
travel_explorer 65,500 N/A Australia Travel

Overview

Custom background script that intercepts network queries and automates navigation on Collabstr to extract creator price lists, niches, and direct booking details without triggering anti-scraping defenses.

Key Features

  • Automated categories navigation & lazy loading parser
  • Direct JSON response capture from hidden API endpoints
  • Extraction of media price tiers (TikTok, Instagram, YouTube)
  • Automated download of portfolio image URLs
  • Rate-limit protection via rotating request signatures

Technologies

Puppeteer Network Hooking XHR Analysis Node.js

Result

Gathers a complete list of influencer costs and contact references dynamically in JSON/CSV formats.

collabstr_output.json
[
  {
    "id": 48291,
    "handle": "kelsey_vlogs",
    "categories": ["lifestyle", "travel"],
    "pricing": {
      "tiktok_video": 350.00,
      "instagram_post": 200.00,
      "youtube_short": 400.00
    },
    "verified": true,
    "location": "Los Angeles, CA"
  },
  {
    "id": 48295,
    "handle": "tech_reviews_daily",
    "categories": ["tech", "gadgets"],
    "pricing": {
      "instagram_story": 150.00,
      "youtube_integration": 1200.00
    },
    "verified": false,
    "location": "Austin, TX"
  }
]

Overview

A multi-threaded document/data extraction process targeting Modash directories to download creator search directories, handles, emails, and demographic breakdowns in high volumes.

Key Features

  • Parallel scraping streams for faster processing
  • Extraction of target creator profile data and verified emails
  • Dynamic scrolling & search pagination handlers
  • Download resume support (avoids duplicate operations)

Technologies

Node.js Stream Axios Network Intercept File System API

Result

Downloads thousands of influencer records containing email addresses within minutes.

Select Run:
Modash_leads_export.xlsx
A B C D E
Creator Name Followers Email Address Platform Follower Engagement
Emily Watson 120,400 emily.growth@gmail.com Instagram 4.2%
Marcus Tech 45,000 m.brody@leadspark.com TikTok 6.8%
DevScale Vlogs 280,000 jchen@devscale.tech YouTube 3.1%
Robert Travel 90,200 robert@travelworld.net Instagram 2.9%
Sarah Cook 510,000 sarah@cookingalbright.com TikTok 5.4%

Overview

A highly modular and resilient e-commerce scraper framework built to scrape product descriptions, pricing configurations, seller details, delivery metrics, and reviews across 10+ wholesale and retail platforms (Amazon, eBay, Etsy, AliExpress, Alibaba, Temu, Walmart, etc.).

Key Features

  • Sub-domain listing and search crawler
  • Multi-marketplace adapters (Amazon, AliExpress, eBay, Etsy, Alibaba, Temu, Walmart, etc.)
  • Rotating residential proxy integration to bypass ip bans
  • Excel/CSV spreadsheets with structured product properties
  • High throughput concurrent download pipelines

Technologies

Python Scrapy Proxy Rotator Playwright Excel Export

Result

Compiles complete store inventory datasets containing thousands of items with live pricing data.

Select Site:
Ecommerce_products_scraped.xlsx
A B C D E
Product Name Price Stock Status Seller Rating Source Site
Apex Gaming Keyboard $89.99 In Stock 4.8 / 5 Amazon
Ergonomic Office Chair $149.50 In Stock 4.5 / 5 eBay
Vintage Ceramic Mug $14.00 Low Stock (2) 5.0 / 5 Etsy
100pc USB-C Cables bulk $45.00 In Stock 4.6 / 5 Alibaba
Smart LED Desk Lamp $22.99 Out of Stock 4.2 / 5 AliExpress

Overview

One-click Google Apps Script engine that scans any Drive folder, your own or a Shared Drive, and crawls recursively through every sub-folder, and builds a complete, clickable index of every single file inside: PDFs, Sheets, Docs, Slides, images, videos, ZIPs and more.

Key Features

  • Recursive sub-folder crawling so nothing gets missed
  • Unlimited scale: pagination handles 10 or 10,000+ files
  • Dual export to a color-coded Google Sheet and a formatted Google Doc
  • Sheet columns: name, type, clickable URL, file ID, parent folder, dates
  • Slack-ready Doc: bold names, blue links, dividers per file
  • Shared Drive + My Drive support, crash-proof silent error logging

Technologies

Google Apps Script Drive API Google Sheets Google Docs

Result

A full color-coded inventory of an entire Drive, with every file named, typed, linked and dated, generated in minutes, ready to share in Slack or audit in Sheets.

drive_scraper · execution log
▶ runAll()  target: "Client Assets" (Shared Drive)
[scan] /Client Assets ................ 214 files
[scan] /Client Assets/Contracts ...... 48 PDFs
[scan] /Client Assets/Reports ........ 122 Sheets
[scan] /Client Assets/Media .......... 1,371 files
[page] 1000 items fetched → requesting next page...
[page] 755 items fetched → done
✔ 2,755 files indexed across 31 sub-folders
✔ Google Sheet created (color-coded by type)
✔ Google Doc created (Slack-ready format)
🔗 Sheet: drive.google.com/spreadsheets/d/1xK...
🔗 Doc:   docs.google.com/document/d/1mQ...
Drive_Index_Report.gsheet
File Name Type URL Parent Folder Modified
Q3_Contract_Final.pdf PDF drive.google.com/file/d/1aB... /Contracts 2026-07-14
Leads_Master_List Google Sheet docs.google.com/spreadsheets/... /Reports 2026-07-21
Onboarding_Guide Google Doc docs.google.com/document/... /HR 2026-06-30
Product_Demo.mp4 Video drive.google.com/file/d/9zX... /Media 2026-07-25
Brand_Assets.zip ZIP drive.google.com/file/d/4fG... /Media 2026-07-08

Overview

A Python command-line tool that turns LinkedIn prospecting into a guided, semi-automated workflow. It runs a "People" search for a target role and region (Medical Device Engineers around Minneapolis in the demo), opens every matching profile in an automated Chrome session, reads the detected experience level and lets the operator accept or reject each lead with a single keystroke.

Key Features

  • Interactive CLI that walks the operator through login, CAPTCHA checks and location-filter setup before the run starts
  • Automated "People" search execution and sequential, profile-by-profile browsing
  • Experience detection with accept / reject prompts, so only qualified profiles are kept
  • Opens each profile in the live browser for instant manual verification when a decision is borderline
  • Real-time console log of every profile visited, the decision taken and the connection status

Technologies

Python Selenium Chrome Automation CLI Workflow Data Parsing

Result

Cuts LinkedIn outreach prep from hours to minutes. Qualified profiles are queued, filtered and opened automatically, so the operator only reviews and connects.

linkedin_leads.py · session log
C:\Users\Apex\Desktop>python linkedin_leads.py
[1/4] Login check ............ OK (session restored)
[2/4] Captcha check .......... none detected
[3/4] ACTION REQUIRED:
      1) Click 'People'
      2) Set LOCATION filter (Minneapolis / nearby)
      3) Apply filter  →  press ENTER when done
[4/4] Search: "Medical Device Engineer"  →  247 results

Opening: linkedin.com/in/sarah-lindqvist-3a7b2
Detected experience: 6 years
Reject? (R = Reject, Enter = Accept):
✅ ACCEPTED by YOU

Opening: linkedin.com/in/daniel-okafor-91c4
Detected experience: 0 years (profile incomplete)
Reject? (R = Reject, Enter = Accept): R
⛔ REJECTED

Opening: linkedin.com/in/priya-raman-eng
Detected experience: 11 years
Reject? (R = Reject, Enter = Accept):
✅ ACCEPTED by YOU
💾 Saved → linkedin_leads.xlsx (row 38)
linkedin_leads.xlsx
A B C D E
Name Headline Location Experience Status
Sarah Lindqvist Medical Device Engineer | R&D Minneapolis, MN 6 yrs Accepted
Daniel Okafor Engineering Student St. Paul, MN 0 yrs Rejected
Priya Raman Sr. Quality Engineer, MedTech Minneapolis, MN 11 yrs Accepted
Marcus Feldt Project Manager | Medical Devices Bloomington, MN 8 yrs Accepted
Elena Vasquez Regulatory Affairs Specialist Eden Prairie, MN 4 yrs Accepted

Overview

A high-throughput Python scraper that spins up parallel worker streams (W1, W2, W3), each driving its own browser to run Google searches for facility and business names, capture the official website or directory listing, and write the result straight back to the source row. Built for healthcare and senior-living directory enrichment where thousands of rows must be resolved quickly.

Key Features

  • Three concurrent worker threads, each with its own browser instance and row range
  • Global CAPTCHA stop: the moment Google challenges one worker, all streams pause, the operator solves it once and the run resumes where it left off
  • Row-level tracking of the search query, matched URL and destination domain (facility sites, medicare.gov listings, etc.)
  • Query-to-row mapping writes results directly into the structured output file
  • Live console showing worker IDs, active rows and saved status for every result

Technologies

Python Selenium Threading Google Search CSV / Sheets

Result

Resolves hundreds of directory rows per hour by running three search streams in parallel, with CAPTCHA handling built in so long runs never silently fail.

scraper.py · W1 / W2 / W3 worker streams
C:\Users\Apex\Desktop>python scraper.py
Workers online: W1 (rows 1–200)  W2 (rows 201–400)  W3 (rows 401–600)

W1 ROW 5
Searching: INDIAN RIVER CENTER WEST MELBOURNE FL
Saved: https://indianriverrehabilitation.com/
--------------------------------------------------
W2 ROW 206
Searching: GLADES HEALTH CARE CENTER PAHOKEE FL
Saved: https://www.medicare.gov/care-compare/details/nursing-home/106018/
--------------------------------------------------
W3 ROW 406
Searching: FOURAKER HILLS REHAB AND NURSING CENTER JACKSONVILLE FL
Saved: https://fourakerhill.com/
--------------------------------------------------
⚠ CAPTCHA detected on W2 → GLOBAL STOP (W1, W3 paused)
   Solve the captcha in the browser, then press ENTER...
▶ Resumed — 3 workers running
--------------------------------------------------
W1 ROW 6
Searching: PARKLANDS CARE CENTER AND REHAB GAINESVILLE FL
Saved: https://parklandscc.com/
--------------------------------------------------
W3 ROW 407
Searching: PALM GARDEN OF SUN CITY SUN CITY CENTER FL
Saved: http://www.palmgardenofsuncitycenter.com/
fl_facilities_resolved.xlsx
A B C D E
Row Facility City / State Matched URL Worker
5 Indian River Center West Melbourne, FL indianriverrehabilitation.com W1
206 Glades Health Care Center Pahokee, FL medicare.gov/care-compare/… W2
406 Fouraker Hills Rehab & Nursing Jacksonville, FL fourakerhill.com W3
6 Parklands Care Center & Rehab Gainesville, FL parklandscc.com W1
407 Palm Garden of Sun City Sun City Center, FL palmgardenofsuncitycenter.com W3
207 Page Rehabilitation & Healthcare Fort Myers, FL pagerehabcenter.com W2

Overview

A Python automation engine that runs the same category search across hundreds of towns and cities on Google Maps, reading the target list from a plain text file (700+ North Carolina cities in the demo), searching each one for categories such as speech therapy and pediatric care, and saving every verified listing to a master spreadsheet.

Key Features

  • City-by-city iteration from simple text lists (nc_cities.txt), so adding a city gets it searched on the next run
  • Live Google Maps automation that captures business title, address, phone, website and sponsored status
  • Exclusion lists (skip_domains.txt) drop unwanted directories and aggregator sites automatically
  • Zero-result cities are logged and skipped, so the run never stalls
  • CAPTCHA prompts pause the run for manual solving and continue afterwards
  • Spreadsheet-ready export: County, City, Category, Provider, Address, Phone, Website and NPI notes

Technologies

Python Selenium Google Maps Text-File Config Multi-Worker

Result

Compiles thousands of localized business and healthcare leads across hundreds of municipalities in hours instead of weeks of manual map searching.

selenium_nc_FINAL.py · city iteration log
C:\Users\Apex\Desktop>python selenium_nc_FINAL.py
Loaded 712 cities from nc_cities.txt  |  14 domains in skip_domains.txt
Category: "speech therapy"

🔵 CITY: Wilkesboro NC
✅ SAVED: "I CAN" Pediatric Therapies
✅ SAVED: NewCastle Speech Services PLLC
🔵 CITY: Fuquay-Varina NC
✅ SAVED: Bound to Bloom Behavioral Institute
✅ SAVED: Sponsored
🔵 CITY: Morrisville NC
✅ SAVED: CareNow Pediatrics
✅ SAVED: Sponsored
🔵 CITY: Rolesville NC
⚠ No valid leads found
🔵 CITY: Zebulon NC
⚠ No valid leads found
🔵 CITY: Wendell NC
⚠ No valid leads found
💾 Progress: 553 rows → providers_nc_all.csv
north_carolina_pediatric_providers.csv
A B C D E
City Provider Category Provider / Practice Phone Website
Raleigh Pediatrics A Plus Family Pediatrics 919-555-1008 Not Found
Charlotte Pediatrics Blythe Blvd Pediatric Group 704-555-8840 blythepeds.com
Wilkesboro Speech Therapy "I CAN" Pediatric Therapies 336-555-1555 icanpediatric.com
Morrisville Pediatrics CareNow Pediatrics 919-555-0182 carenowpeds.com
Hickory Pediatrics Advanced Rehab & Pediatrics 704-555-0000 Not Found
Fayetteville Pediatrics Ramsey St Pediatric Clinic 910-555-7337 ramseypeds.com

Overview

A post-processing workflow that turns a raw, scraped lead list into a clean master database. A pandas script (work.py) deduplicates the CSV, the cleaned file is uploaded to Google Sheets, and a custom conditional-formatting rule (=COUNTIF(G:G, G2)>1) instantly flags any phone number that still appears more than once.

Key Features

  • One-command deduplication from the Windows command prompt, ending with "DONE: leads_deduplicated.csv created successfully"
  • Cleaned CSV imported straight into Google Sheets for team-wide visibility
  • Custom formula-based conditional formatting colour-codes duplicate phone numbers across the entire column
  • Handles 1,700+ rows of County, City, Provider Category, Practice, Address, Email, Phone, Website and Notes in one pass
  • Repeatable: re-run the script on every new scrape and the sheet stays clean

Technologies

Python pandas CSV Google Sheets Conditional Formatting

Result

Messy raw scrapes become deduplicated, visually audited master lists ready for outreach or CRM import, with no manual row-by-row checking.

work.py · deduplication run
C:\Users\Apex\Desktop>python work.py
Loaded  : leads_raw.csv  (1,894 rows)
Dedup on: Phone Number + Provider Name
Removed : 134 exact duplicates
✅ DONE: leads_deduplicated.csv created successfully (1,760 rows)

Google Sheets → File → Import → leads_deduplicated.csv
Format → Conditional formatting
  Range : G2:G1760
  Rule  : Custom formula   =COUNTIF(G:G, G2)>1
  Style : Red fill
Result: 12 phone numbers still flagged for manual review
leads_deduplicated.csv
A B C D E
County City Provider / Practice Phone Number Website
Alamance County Burlington Burlington Pediatrics (336) 555-8316 burlingtonpeds.com
Alamance County Burlington KidzCare Pediatrics (336) 555-7337 kidzcarepeds.com
Alamance County Graham Small World Therapy (919) 555-5437 smallworldtherapy.com
Alamance County Elon Kernodle Clinic (336) 555-2314 Not Found
Cumberland County Fayetteville Rainbow Pediatrics (919) 555-5437 rainbowpeds.com
Cumberland County Fayetteville Cape Fear Valley Pediatrics (910) 555-6431 capefearvalley.com

Overview

A monitoring script (sheet_watcher.py) that attaches to an already-open Chrome tab through the remote-debugging port, watches a live Google Sheet such as a monthly Water Leakage Report, and raises an audible alarm the instant a new row or status change appears, so facility teams react to a leak in seconds instead of at the next manual check.

Key Features

  • Attaches to the operator's existing Chrome session, with no separate login or API keys required
  • Polls the active sheet tab on a fixed interval and reloads the page automatically
  • Detects new rows, edited status cells and missing meter readings against a stored baseline
  • Audible + console alarm ("NEW DATA ADDED – ALARM STARTED") that stops automatically once the change is cleared
  • Timestamped terminal log of every row count and state change for later review

Technologies

Python Chrome DevTools Protocol Selenium Google Sheets Alerting

Result

Operations teams are notified of critical spreadsheet updates the moment they happen, drastically reducing response time for leak reports and other shared trackers.

sheet_watcher.py · live monitor
C:\Users\Apex\Desktop>python sheet_watcher.py
[03:08:12] 🚀 Attaching to your Chrome (debug mode)...
[03:08:13] ✅ Attached to your Chrome browser
[03:08:15] 📊 Watching "January Water Leakage Report" — Initial rows: 16
[03:08:20] 🚨 NEW DATA ADDED – ALARM STARTED   (rows: 17)
[03:08:33] 🔄 Page reloaded in your browser
[03:08:36] 🛑 DATA REMOVED – ALARM STOPPED     (rows: 16)
[03:08:44] 🚨 NEW DATA ADDED – ALARM STARTED   (rows: 17)
[03:08:48] 🔄 Page reloaded in your browser
[03:08:56] 🛑 DATA REMOVED – ALARM STOPPED
[03:09:06] 🔄 Page reloaded in your browser — no change
[03:09:36] 🔄 Page reloaded in your browser — no change
January Water Leakage Report · 01-16-2026 Analysis
A B C D E
Building Current Reading Previous Reading Usage (gal) Status
330 12,480 No Reading
336 14,910 11,020 3,890 Leak
300 5,140 5,090 50 OK
76 7,570 7,510 60 OK
324 12,860 9,140 3,720 Leak
3984 16,220 15,400 820 Possible Leak
2604 3,610 3,580 30 OK

Overview

A lightweight desktop automation script (auto_scroll_ctrl.py) that keeps large lead directories and Google Sheets moving on their own, scrolling on a fixed interval and sending periodic keystrokes so long tables load, render and can be parsed end to end without anyone touching the mouse.

Key Features

  • Configurable scroll interval, scroll amount and key-press interval in a three-line settings block
  • PyAutoGUI fail-safe: move the mouse to a screen corner and the script stops instantly
  • Timestamped log of every scroll and key event for auditability
  • Works on any scrolling surface, from web directories to Google Sheets and internal dashboards
  • Pairs with the lead tracker sheet (City, Provider Category, Practice, Address, Email, Phone, Website) so thousands of rows can be reviewed hands-free

Technologies

Python PyAutoGUI Google Sheets Command Prompt

Result

Makes reviewing and validating thousands of directory rows a hands-free process. The script does the scrolling and the operator only reads.

auto_scroll_ctrl.py · execution log
C:\Users\Apex\Desktop>python auto_scroll_ctrl.py
=== AUTO SCROLL + CTRL SCRIPT STARTED ===
Move mouse to screen corner to STOP (fail-safe ON)
SCROLL_INTERVAL = 2s   CTRL_INTERVAL = 10s   SCROLL_AMOUNT = -120

[07:07:02] Scrolled down
[07:07:04] Scrolled down
[07:07:06] Scrolled down
[07:07:08] Scrolled down
[07:07:10] Scrolled down
[07:07:10] CTRL pressed
[07:07:12] Scrolled down
[07:07:14] Scrolled down
[07:07:16] Scrolled down
[07:07:18] Scrolled down
[07:07:20] Scrolled down
[07:07:20] CTRL pressed
...
[07:12:40] Fail-safe triggered — script stopped (168 scrolls, 33 key presses)
Tracker · Leads
A B C D E
City Provider Category Provider / Practice Name Phone Number Email
Burlington Pediatric Speech Therapy New Discovery Speech & Language (678) 555-0108 Not Found
Burlington Pediatrician Arrowhead Blvd Pediatrics (919) 555-0202 Not Found
Graham Pediatric Occupational Therapy Small World Therapy – Graham (919) 555-5437 admin@smallworldtherapy.com
Mebane Pediatric Speech Therapy Speech Stars (919) 555-8480 admin@speechstarsnc.com
Haw River ABA Center Proud Moments ABA (336) 555-7655 info@proudmomentsaba.com
Ossipee Pediatric Speech Therapy Take Back Speech Therapy (336) 555-6567 Not Found

Overview

An automated bot that searches Google Maps for a service category in a target area (pressure washing and exterior cleaning companies in the demo), opens each listing, and harvests the business name, website, physical address, phone number and email into a structured lead file.

Key Features

  • Automated search, result scrolling and listing-by-listing navigation on Google Maps
  • Extracts name, website, address, phone and email; missing fields are logged as "Not found" instead of breaking the run
  • Progress counter ([75/119] Processing) with a saved-entry confirmation for every business
  • Handles dynamic map loading and lazy-rendered detail panels
  • Runs in a visible or headless browser

Technologies

Python Selenium Google Maps Chrome Automation CSV Export

Result

Builds a complete local business directory with enriched contact data in minutes. Hours of manual copy-paste are replaced by a single command.

maps_scraper.py · "pressure washing Gainesville FL"
C:\Users\Apex\Desktop>python maps_scraper.py
Query: "pressure washing Gainesville FL"  →  119 listings found

➡ [75/119] Processing
   ✅ Saved: Sunshine Pressure Pros
   📍 Gainesville, FL
   📞 +1 352-555-0148
   🌐 Not found
   📧 Not found

➡ [77/119] Processing
   ✅ Saved: Marsh Window Cleaning & Power Washing
   📞 +1 352-555-0390
   🌐 http://gainesvillepressurewashing.com/
   📧 Not found

➡ [84/119] Processing
   ✅ Saved: Dirty Dog Power Wash Supply
   📍 8260 SE 58th Ave UNIT 3, Ocala, FL 34480
   📞 +1 352-555-0124
   🌐 http://dirtydogpws.com/
   📧 sales@dirtydogpws.com

➡ [85/119] Processing ...
gainesville_pressure_washing.csv
A B C D E
Business Name Phone Website Email Address
Sunshine Pressure Pros +1 352-555-0148 Not found Not found Gainesville, FL
Marsh Window Cleaning & Power Washing +1 352-555-0390 gainesvillepressurewashing.com Not found Gainesville, FL
Dirty Dog Power Wash Supply +1 352-555-0124 dirtydogpws.com sales@dirtydogpws.com 8260 SE 58th Ave, Ocala, FL
Saxco Solutions LLC +1 352-555-0219 saxcosolutions.com info@saxcosolutions.com Alachua, FL
Gator Clean Exteriors +1 352-555-0377 gatorcleanfl.com Not found Newberry, FL

Overview

A login-protected portal scraper (ascend_scraper.py) that authenticates into the Ascend reseller portal, loads the full publication pricing catalogue, and walks through every listing, saving each unique publication with its niche, pricing and attributes while skipping records that were already captured.

Key Features

  • Automated login sequence and session management
  • Dynamic loading and parsing of the publication database (1,462+ catalogue entries in the demo)
  • Real-time duplicate detection, with "Duplicate skip" or "Saved" logged for every record
  • Works with the portal's own search filters, niche tags and paginated tables
  • Visible or headless Chrome execution

Technologies

Python Selenium Chrome Automation Session Management CSV Export

Result

Keeps a complete, duplicate-free copy of the reseller catalogue up to date without any manual data entry.

ascend_scraper.py · catalogue sync
C:\Users\Apex\Desktop>python ascend_scraper.py
🔐 Logging in to Ascend Reseller Portal ... OK
📚 Publications loaded: 1,462   |   already saved: 918

⏩ Duplicate skip: Business Insider
⏩ Duplicate skip: Entertainment Monthly News
⏩ Duplicate skip: Influencer Daily
💾 Saved: Northampton Herald
💾 Saved: Yardley Voice
⏩ Duplicate skip: Pulse Sports
💾 Saved: City Sun Times
💾 Saved: Essex Reporter
💾 Saved: Colchester Sun
⏩ Duplicate skip: Retail Insider
💾 Saved: The Chaffee County Times
💾 Saved: Warwick Journal
💾 Saved: Newton Gazette
💾 Saved: Mid East Info
⏩ Duplicate skip: Notion Online
💾 Saved: Pagosa Springs Sun
...
✅ Run complete — 544 new publications saved, 918 duplicates skipped
ascend_publications.csv
A B C D E
Publication Niche Price DA Status
Northampton Herald News / Local $150 42 Saved
Yardley Voice News / Local $150 38 Saved
City Sun Times News / Lifestyle $150 45 Saved
Business Insider Business $3,500 94 Duplicate skip
Essex Reporter News / Local $150 40 Saved
Retail Insider Business / Retail $220 51 Duplicate skip

Overview

A watcher script that connects to a running Chrome instance over the remote-debugging port, refreshes the YouTube Subscriptions feed on a loop, stores a baseline of what is already there, and fires a loud, continuous alarm the moment a new upload appears.

Key Features

  • Attaches to the operator's logged-in Chrome (--remote-debugging-port=9222), with no API quota and no separate login
  • Automated refresh cycle on youtube.com/feed/subscriptions with baseline comparison
  • Console status for every cycle: Refreshing page…, Baseline stored, NEW VIDEO DETECTED !!!
  • Keyboard controls while running: Q stops the alarm and X exits the watcher
  • Infinite alarm until acknowledged, so a new video is never missed

Technologies

Python Chrome DevTools Protocol Selenium YouTube CLI

Result

Eliminates manual feed checking. The operator is alerted within seconds of a new upload from any monitored channel.

yt_watcher.py · subscription feed monitor
C:\Users\Apex\Desktop>python yt_watcher.py
🚀 Connected to Chrome (DEBUG MODE)
📺 Watching YouTube SUBSCRIPTIONS page
🔔 LOUD infinite alarm on new video
🛑 Press Q = stop alarm   |   ❌ Press X = exit

🔄 Refreshing page...
ℹ️ Baseline stored (24 videos)
🔄 Refreshing page...
✅ No new video
🔄 Refreshing page...
✅ No new video
🔄 Refreshing page...
🚨🚨 NEW VIDEO DETECTED !!! 🚨🚨
🛑 Q pressed → Alarm STOPPED (watcher running)
🔄 Refreshing page...
✅ No new video
🔄 Refreshing page...
🚨🚨 NEW VIDEO DETECTED !!! 🚨🚨
yt_watcher · detections.csv
A B C D E
Time Channel Video Title Uploaded Action
09:20:41 Forbes Breaking News 'Is That Normal Operating Procedure?' … 5 minutes ago Alarm → Q
09:21:07 Ghost Night January 17, 2026 1 minute ago Alarm → Q
09:21:36 No new video
09:22:04 Forbes Breaking News 'The Last Thing We Need Is More Weapons' … just now Alarm → Q

Overview

A verification assistant for vehicle inventory audits. The script reads each row of a master Google Sheet (Year, Make, Model, VIN, sticker image URL), opens the door-jamb sticker photo in an automated Chrome window, and asks the operator to confirm the match with a single keystroke, writing the result back to the sheet as it goes.

Key Features

  • Row-by-row iteration through the master spreadsheet with the current row shown in the terminal (Row 159, Row 160 …)
  • Opens each sticker image URL automatically in Chrome for instant visual inspection
  • Interactive prompt to confirm, reject or quit, with "Saved" written back to a status column
  • Keeps Year, Make, Model, VIN and Image URL in sync with the sheet
  • Resume from any row number, so long audits can be split across sessions

Technologies

Python Selenium Google Sheets API Chrome Automation CLI

Result

Turns a tedious visual VIN audit into a rapid keystroke workflow, cutting data-entry errors while keeping a human in the loop for every match.

vin_verify.py · audit session
C:\Users\Apex\Desktop>python vin_verify.py --start-row 159
📌 Row 159
VIN (Sheet): 1G1FB1RS1G0152595
Opening URL: https://smartauctionprod.blob.core.windows.net/images/4615900-….jpeg
Enter y / n / q: y
💾 Saved

📌 Row 160
VIN (Sheet): 1GCPTFEK4R1149733
Opening URL: https://smartauctionprod.blob.core.windows.net/images/4616000-….jpeg
Enter y / n / q: y
💾 Saved

📌 Row 161
VIN (Sheet): 3GNAXNEG6PL125236
Opening URL: https://smartauctionprod.blob.core.windows.net/images/4616100-….jpeg
Enter y / n / q: n
⚠ Mismatch logged — flagged for review

📌 Row 162
VIN (Sheet): 1C4HJXCN8PW614608
Opening URL: https://smartauctionprod.blob.core.windows.net/images/4616200-….jpeg
Enter y / n / q: y
💾 Saved
vehicle_inventory_audit.gsheet
A B C D E
Year Make Model VIN Verified
2016 Chevrolet Cruze 1G1FB1RS1G0152595 Saved
2024 Chevrolet Silverado 1500 1GCPTFEK4R1149733 Saved
2023 Chevrolet Equinox 3GNAXNEG6PL125236 Mismatch
2023 Jeep Wrangler 1C4HJXCN8PW614608 Saved
2018 GMC Sierra 3500 1GD32WEY1JF263262 Saved
2023 Cadillac CT5 1G6DS5RK6P0106656 Pending

Overview

A Python workflow (script.py) that scans an input folder of PDFs such as utility bills, invoices, cheques and accounting notices, reads each file, classifies it into a category such as Utilities, Checks, HAP Adjustment or Ledger Cleanup, renames it to a consistent pattern and writes a tracking row (file name, Slack channel, tagged person, category) to CSV and Google Sheets.

Key Features

  • Batch directory scanning with per-file progress ("Processing: scan (18).pdf")
  • Text-based classification rules that map document content to business categories
  • Standardised renaming (e.g., 19440 ELMWOOD ST_Utilities_2026.01.pdf) for clean archiving
  • Tracking output with Slack channel and @person tags, ready for team routing
  • Graceful error handling: malformed or unreadable PDFs are logged and skipped, never crash the run

Technologies

Python PyMuPDF CSV Google Sheets VS Code

Result

Heavy volumes of incoming paperwork are sorted, renamed and catalogued into a shared tracker with zero manual data entry.

script.py · MailAutomation terminal
PS C:\Users\Apex\Desktop\MailAutomation> python script.py
Scanning input_pdfs/ ... 20 files found

Processing: scan (16).pdf
✅ Saved: 19440 ELMWOOD ST_Utilities_2026.01.pdf
Processing: scan (17).pdf
✅ Saved: 11397 WHITCOMB ST_Utilities_2026.01.pdf
Processing: scan (18).pdf
✅ Saved: 20164 TERRELL_Other_2026.03.pdf
Processing: scan (19).pdf
MuPDF error: format error: object is not a stream
⚠ Skipped → moved to review/
Processing: scan (2).pdf
✅ Saved: 11285 COURVILLE_Other_2026.03.pdf
Processing: scan (20).pdf
✅ Saved: 12265 LONGVIEW_Other_2026.03.pdf

Done: 19 classified, 1 sent to review  →  output.csv updated
output.csv
A B C D
File Name Slack Channel Tag Person Category
19440 ELMWOOD ST_Utilities_2026.01.pdf mail @Amanda Utilities
11397 WHITCOMB ST_Utilities_2026.01.pdf mail @Amanda Utilities
8181 HOUSE ST_Utilities_2026.03.pdf mail @Amanda Utilities
625 KENMOOR_Checks_2026.03.pdf accounting @here Checks
20164 TERRELL_Other_2026.03.pdf mail @Amanda HAP Adjustment
scan (19).pdf Review

Overview

A cleanup script (clean_leads.py) that standardises messy lead exports with repeated header rows, inconsistent phone and fax formats and empty "Preferred Contact" fields, then produces a single clean CSV that imports straight into Google Sheets for collaborative call tracking.

Key Features

  • One-command execution from the Windows command prompt with a clear completion message ("DONE: Fax & Preferred Contact fixed. All leads kept.")
  • Removes duplicate header blocks and normalises phone / fax formatting
  • Keeps every lead, because cleaning never silently drops rows
  • Direct import of clean_leads.csv into Google Sheets (separator auto-detected)
  • Structured columns: County, Category, Name, Address, Phone, Fax, Email, Spoke To, Notes

Technologies

Python CSV Google Sheets Command Prompt Data Normalisation

Result

Unstructured raw contact lists become clean, filterable Google Sheets that outreach teams can start calling from immediately.

clean_leads.py · cleanup run
C:\Users\Apex\Desktop>python clean_leads.py
Input : leads_raw.csv            (412 rows, 9 columns)
Found : 17 repeated header rows  → removed
Fixed : 96 phone numbers, 31 fax numbers normalised to ###-###-####
Fixed : 58 empty "Preferred Contact" cells → defaulted to Phone
✅ DONE: Fax & Preferred Contact fixed. All leads kept.
Output: clean_leads.csv          (395 rows)

Google Sheets → File → Import → clean_leads.csv
  Import location : Create new spreadsheet
  Separator type  : Detect automatically
✅ Import complete — sheet "Final" ready
clean_leads.csv
A B C D E
County Provider Name Phone Email Notes
Pender County Black River Family Medicine 910-555-5721 Not Found 01/05: On hold for an excessive amount of time.
Pender County Burgaw Medical Center 910-555-3377 Not Found 01/05: Left VM.
Pender County Wilmington Health 910-555-6558 billing@wilmingtonhealth.com 01/05: Sent email.
Pender County Mill Creek Family Practice 910-555-2515 millcreekfp@gmail.com 01/05: Will respond to email or call if interested.
Tyrrell County Columbia Medical Center 252-555-0689 Not Found 01/05: Line busy, unable to leave voicemail.
Pender County Coastal Carolina Pediatrics 910-555-7307 Not Found 01/05: Transferred to referral coordinator VM.

Overview

An OCR-driven sorting pipeline that reads every scanned PDF in an input_pdfs folder, extracts text with PyMuPDF and Tesseract (fast OCR on the top region of the page), decides where the document belongs and files it automatically. Parsed documents go into date-based output folders (output/2026.03), anything uncertain goes to review/, and every action is logged to output.csv.

Key Features

  • Batch processing with per-file terminal progress (Processing: scan (1).pdf, scan (2).pdf …)
  • Hybrid extraction: native PDF text first, Tesseract OCR fallback for image-only scans
  • Conditional routing into output/<year.month>/ or review/ using shutil
  • Standardised file renaming with account number, name, category and period
  • Tabular output.csv summary of every processed file for record-keeping

Technologies

Python PyMuPDF (fitz) pytesseract Pillow CSV

Result

Bulk scanned paperwork is digitised, sorted and archived automatically. Manual sorting time drops to near zero while records stay structured.

script.py · OCR sorting run
PS C:\Users\Apex\Desktop\MailAutomation> python script.py
Tesseract: C:\Program Files\Tesseract-OCR\tesseract.exe
INPUT_FOLDER = input_pdfs   OUTPUT_FOLDER = output   REVIEW_FOLDER = review

Processing: scan (1).pdf
✅ Saved: 12271 WYOMING_Other_2026.03.pdf
Processing: scan (2).pdf
✅ Saved: 11285 COURVILLE_Other_2026.03.pdf
Processing: scan (3).pdf
✅ Saved: 11406 STOCKWELL_Other_2026.03.pdf
Processing: scan (4).pdf
✅ Saved: 10498 MERLIN_Other_2026.03.pdf
Processing: scan (5).pdf
✅ Saved: 16225 MANNING_Other_2026.03.pdf
Processing: scan (6).pdf
✅ Saved: 16236 APPOLINE_Other_2026.03.pdf
Processing: scan (7).pdf
⚠ Low OCR confidence → review/scan (7).pdf

Done: 19 sorted into output/2026.03, 1 sent to review  →  output.csv
MailAutomation/ · after run
📁 input_pdfs/ 20 scans (source)
📁 output/
📁 2026.03/ 19 files
📄 11406 STOCKWELL_Other_2026.03.pdf
📄 10498 MERLIN_Other_2026.03.pdf
📄 16225 MANNING_Other_2026.03.pdf
📄 16236 APPOLINE_Other_2026.03.pdf
📁 review/ 1 file flagged
📄 scan (7).pdf low OCR confidence
📊 output.csv 20 rows · file, category, destination
🐍 script.py

Overview

A locally hosted Lead Generation System (v8.1) with a full control panel: create multiple scraping sessions, enter keywords and locations, set speed and safety levels, choose which fields and social links to capture, and watch results arrive in a live table with one-click CSV / Excel export.

Key Features

  • Multi-session management, so "electricians in Chicago" and "plumbers in New York" run side by side, each with its own Start, Stop and Reset
  • Keyword + location + max-results configuration with "skip first N" for resuming interrupted runs
  • Speed presets from Normal to Ultra Fast with a fine delay multiplier, CAPTCHA pause and anti-bot safety controls
  • Quality filters (require phone / email, prefer website data, min rating and reviews) and deduplication by name, phone, website or address
  • Social link extraction: Facebook, Instagram, LinkedIn, Twitter/X, YouTube, TikTok
  • Live analytics (Saved, Skipped, Current Listing, Loaded Listings), session logs and CSV / Excel / Copy export

Technologies

React Node.js Puppeteer Google Maps Excel / CSV Export

Result

A self-serve lead engine. Anyone on the team can launch a targeted local-business scrape and download a clean, deduplicated dataset without touching code.

Lead Generation System v8.1 · Analytics & Logs
Session 3  (sess_1776127322214_b7jql)   status: Running
keyword: "electricians in Chicago"   location: New York   max: 50
speed: 🔥 Ultra Fast (x1.0)   pause on CAPTCHA: on   socials: LinkedIn, Twitter/X, YouTube, TikTok

[05:43:12] Session started
[05:43:14] Loaded listings: 20
[05:43:19] ✅ Saved  #1  Windy City Electric Co.        (phone ✓ email ✓ website ✓)
[05:43:23] ✅ Saved  #2  Lakeview Electrical Services    (phone ✓ email ✗ website ✓)
[05:43:26] ⏭ Skipped     Ace Handyman Pros              (no phone)
[05:43:31] ✅ Saved  #3  North Shore Electricians        (phone ✓ email ✓ website ✓)
[05:43:35] 🔗 Socials    facebook, instagram, linkedin found
[05:43:40] Loaded listings: 40
[05:43:44] ✅ Saved  #4  Bright Spark Electric           (phone ✓ email ✓ website ✗)
...
Saved Leads: 8   Skipped Leads: 3   Current Listing: 12   Loaded Listings: 40
session3_electricians_chicago.xlsx
A B C D E
Business Phone Email Website Socials
Windy City Electric Co. (312) 555-0142 office@windycityelectric.com windycityelectric.com FB · IG · LI
Lakeview Electrical Services (773) 555-0198 lakeviewelectrical.com FB
North Shore Electricians (847) 555-0117 hello@northshoreelec.com northshoreelec.com FB · IG · LI · YT
Bright Spark Electric (312) 555-0163 info@brightsparkchi.com IG · TikTok
Loop Power & Lighting (312) 555-0130 looppower.com LI · X

Overview

A scraper for directory tables where key attributes are shown only as icons for image allowed, sponsored, indexed, do-follow, niche flags and price multipliers rather than text. The script walks 1,200+ listings, reads each icon's state from the DOM and converts it into clean Y / blank / "x2.0 cost" values alongside the visible fields (genres, price, DA, DR, turnaround, region).

Key Features

  • Icon-state detection (active vs inactive) converted to explicit text flags per column
  • Captures genres, price, DA / DR metrics, turnaround time, region and staff / niche indicators
  • Side-by-side terminal progress, such as "161/1205 Saved: Ritz Herald News ['', 'Y', 'Y', 'Y', 'Y']"
  • Handles pagination and lazy-loaded rows across the full directory
  • Writes a structured local dataset ready for filtering and pricing analysis

Technologies

Python Selenium / Playwright DOM Parsing CSV Export Windows Terminal

Result

Hundreds of icon-only attributes become searchable text data in one run, with no manual copy-paste and no guesswork about what each icon meant.

icons_scraper.py · 1,205 listings
C:\Users\Apex\Desktop>python icons_scraper.py
Columns: [Image, Sponsored, Indexed, DoFollow, Niches/Cost]

152/1205 Saved: WW Journals ['Y', 'Y', 'Y', 'Y', 'x2.0 cost']
153/1205 Saved: Dollar Thinking ['', '', '', '', '']
154/1205 Saved: ABC Money ['', '', '', 'Y', '']
155/1205 Saved: Healthcare Business Today ['', 'Y', '', '', '']
156/1205 Saved: Sports Tech Today Staff ['Y', 'Y', 'Y', 'Y', 'Y']
158/1205 Saved: Nohoarts District ['Y', 'Y', 'Y', 'Y', 'Y']
160/1205 Saved: Coin Review News Staff ['Y', 'Y', 'Y', 'Y', 'Y']
161/1205 Saved: Ritz Herald News ['', 'Y', 'Y', 'Y', 'Y']
162/1205 Saved: Mississippi Independent ['Y', 'Y', 'Y', 'Y', 'x2.0 cost']
163/1205 Saved: Meditech Today Staff ['Y', 'Y', 'Y', 'Y', 'x3.0 cost']
165/1205 Saved: PC World Solutions ['', '', '', '', '']
167/1205 Saved: Im Techies ['', '', '', '', '']
...
💾 Checkpoint written: publications_icons.csv (167 rows)
publications_icons.csv
A B C D E F
Publication Genres Price DA / DR TAT Flags (Img · Spons · Idx · DoFollow)
WW Journals News $150 28 / 45 3-5 Days Y · Y · Y · Y · x2.0 cost
Sports Tech Today News, Sports $150 56 / 41 1-3 Days Y · Y · Y · Y · Y
Ritz Herald News News $150 42 / 45 3-5 Days – · Y · Y · Y · Y
Dollar Thinking Finance $150 24 / 27 3-5 Days – · – · – · – · –
Meditech Today Health $150 48 / 58 3-5 Days Y · Y · Y · Y · x3.0 cost

Overview

A locally hosted, multi-session video studio. Paste a YouTube, TikTok or direct link (or drop a file), and VideoStudio ingests the media, lets you cut precise clips by timestamp, transcribes them with Whisper, extracts audio and merges selected clips into a single MP4 or MP3, all from one dashboard.

Key Features

  • Multi-session workspaces (Session 1, Session 2 …) so several projects can be prepared in parallel
  • Media ingestion from YouTube, Twitter/X, TikTok, direct URLs or local MP4 / MKV / MOV / MP3 / WAV files
  • Clip cutting with start / end selectors, a max-duration guard and per-clip preview
  • Whisper speech-to-text for a full video or an individual clip; Extract Audio (MP3) in one click
  • Merge engine: reorder clips and export as one MP4 video or MP3 audio file

Technologies

Node.js Next.js / React Whisper FFmpeg Media Download API

Result

A complete media pipeline for creators and agencies. Raw links become trimmed, transcribed and merged clips in minutes.

session_1/export_manifest.json
{
  "session": "Session 1",
  "source": {
    "title": "STOP Building AI Agents (Most People Are Doing It Wrong)",
    "type": "youtube",
    "duration": "22:28"
  },
  "clips": [
    { "id": 1, "start": "00:00", "end": "00:30", "transcribed": true,
      "transcript": "Most people build agents backwards. They start with..." },
    { "id": 2, "start": "04:12", "end": "06:40", "transcribed": true,
      "transcript": "The mistake is treating the model as the product..." }
  ],
  "audio_extracted": "session_1/clip_1.mp3",
  "merge": {
    "order": [2, 1],
    "output": "session_1/merged_output.mp4",
    "status": "done"
  }
}

Overview

A contact-enrichment tool (email_finder.py) that takes an investor spreadsheet of VC firms, accelerators and angel groups, visits every website, reads the full HTML source plus contact and about pages, and writes back every business email it can find, or a clear status note when it can't.

Key Features

  • Batch processing of thousands of target websites from an input CSV / Google Sheet
  • Deep email discovery: full page source, contact, privacy and about pages and linked sub-pages, returning all emails found rather than just the first
  • Status notes written back per row: "No email found", "Timeout", "Site not reachable"
  • Auto-save checkpoints every 10 rows ("Saved till row 40") so an interrupted run resumes exactly where it stopped
  • Multi-threaded requests with per-site timeouts to keep long runs moving
  • Export to Google Sheets or CSV with the original investor columns preserved

Technologies

Python Requests BeautifulSoup Multi-threading Google Sheets / CSV

Result

Raw investor directories become outreach-ready lead databases with verified contact emails. Hours of manual site checking are replaced by an unattended run.

email_finder.py · investors_with_emails
C:\Users\Apex\Desktop>python email_finder.py
Input: investors.csv (1,240 rows)   threads: 4   timeout: 10s

Checking (36): https://www.greenhousecap.com
Checking (37): https://www.greenhouse.ventures/
Checking (38): http://greentownlabs.com
Checking (39): https://www.greysiloventures.com
Checking (40): https://www.greycroft.com
💾 Saved till row 40
Checking (41): https://www.greymattercapital.com
Checking (42): http://www.gridventures.com
Checking (43): https://www.grid110.org/
Checking (44): https://griegkapital.no/
Checking (45): https://www.griffinaccelerator.com.au
Checking (46): http://www.grimgold.se
Checking (47): https://www.grindcapital.org
Checking (48): http://gringottsventures.com
Checking (49): https://gritrd.com/
Checking (50): https://www.gritlabs.io
💾 Saved till row 50
Checking (51): https://www.grix.vc/
investors_with_emails.gsheet
A B C D E F
Investor Name Website Investor Type Region Extracted Email Status / Notes
GO Ventures goventures.com.mt Corporate VC EU talktous@goventures.com.mt Email Verified
Gold House goldhouse.org Accelerator USA contact@goldhouse.org Email Verified
Golden Seeds goldenseeds.com Angel Group USA info@goldenseeds.com Email Verified
Golden Gate VC goldengate.vc VC East Asia N/A No email found
Graph Paper Capital graphpapercapital.xyz VC USA team@graphpapercapital.xyz Email Verified
Gravity Fund gravityfund.vc VC USA N/A Timeout

Overview

A research and extraction tool for tax-sale and property-auction investors. Running as an overlay panel directly on the USTLA research listings (uls.ustla.com), it pulls every auction listing plus its detail page into a clean, multi-tab Excel workbook organised by month, with county, state, sale date, auction type, venue, opening bids, parcel IDs and property URLs.

Key Features

  • One-click list + detail scraping with filters for auction type (Deed, Redeemable Deed, Lien), status and sale-date range, or all pages at once
  • Floating in-browser panel with live counters: auctions, details, failed, 429 hits, mismatches and saved
  • Retry-failed and resume controls that survive rate limits and page reloads
  • Deep field capture: county, state, sale date & time, post date, auction type, sale venue, contact, opening bid, estimated value, parcel / APN and direct USTLA URL
  • Excel export (USTLA_Auctions_Filled.xlsx) with one tab per month (Sep 2025, Oct 2025 …), styled headers and auto-fitted columns, ready for pivot tables

Technologies

JavaScript Browser Console / DevTools DOM Parsing XLSX Export Excel

Result

Thousands of tax deed and lien records across multiple states are collected and underwriting-ready in minutes, eliminating manual data entry from the acquisition pipeline.

USTLA Scraper · overlay panel + console
[ULS] Ready. Check settings ▶ press Start. Filters are set in this panel.
Auction Type : ☑ Deed  ☑ Redeemable Deed  ☐ Lien
Status       : Current + Future & Past      Min Sale Date: 09/01/2025
Pages        : all (48)                     Displaying 1–100 of 4,796 auctions

▶ Start
[list]   page 1/48 ...... 100 rows
[list]   page 2/48 ...... 100 rows
[detail] 1–100 ........... 100 ok   0 failed
[detail] 101–200 ......... 99 ok    1 failed  (429 → retry queued)
[list]   page 48/48 ..... 96 rows
[detail] retry failed .... 1 ok

Auctions: 4,796   Details: 4,796   Failed: 0   429 hits: 3   Mismatch: 0
💾 Saved: USTLA_Auctions_Filled.xlsx  (tabs: Progress, Sep 2025 … Feb 2026)
⬇ Download
USTLA_Auctions_Filled.xlsx · Sep 2025
A B C D E F
County State Sale Date Auction Type Sale Venue Opening Bid
Dallas County Texas 2025-09-02 09:00 AM Tax Deed Online auction $14,200
Harris County Texas 2025-09-02 10:00 AM Tax Deed Bayou City Event Center, Houston $9,850
Fort Bend County Texas 2025-09-02 10:00 AM Sheriff Sale County Fairgrounds – Building C $22,400
Denton County Texas 2025-09-02 10:00 AM Tax Deed Online auction $7,300
Hill County Texas 2025-09-02 10:00 AM Tax Certificate Courthouse Steps, Hillsboro $3,120
El Paso County Texas 2025-09-02 10:00 AM Tax Deed Online auction $11,675

Overview

A complete operations system rather than a single script, built for an independent construction-consulting firm and documented in a 27-page System Manual. ClickUp is the operational board, Google Drive and Apps Script are the backend (folder tree, workbook, intake forms and release checks), Zapier carries events between platforms, Smartlead and Prospeo run outbound lead generation, HubSpot holds the CRM record and Dropbox is the client-facing file layer. Every engagement moves through one board in one direction, with hard approval gates and an audit log at every step.

Key Features

  • Unified ClickUp operational board: every engagement tracked from New Opportunity → Intake → Scope Development → Owner Approval → Proposal → Agreement & Deposit → Field Planning → Field Release
  • Google Apps Script backend engine: one project provisions the whole Drive tree (00 Backend, 01 Templates, 02 Forms, 03 Engagements), the master workbook and every intake form, validates responses and runs the fieldReleaseCheck audit
  • Zapier multi-bridge layer: one intake submission fans out to four Zaps at once (ClickUp task, Drive folder, Dropbox folder, Engagement Register) and every status change writes a row to the Gate Log
  • Automated outbound & lead generation: city-by-city Smartlead campaigns on authenticated, warmed mailboxes with suppression checks and Smart Dialer, plus Prospeo lead sourcing synced into HubSpot
  • Structured document lifecycle: per-engagement folders, templated Docs / Sheets, tiered intake modules (Tier 1–3), Site Contact forms and gated proposal / legal release documents
  • Fail-safe gate control & audit trail: no stage can be skipped, every backward move is a block, nothing is released without a complete site-contact set and a human sign-off, and the Gate Log's "Evidence confirmed" and "Approved by" columns are never auto-filled
  • Operator SOP manual: who does what, daily operation, exceptions, timestamps and do-it-yourself runbooks so the client's own team runs the system without us

Technologies

ClickUp Google Apps Script Zapier Smartlead.ai Prospeo HubSpot Google Drive

Result

Eliminates manual hand-offs across eight platforms, cuts administrative overhead by over 80% and guarantees that every proposal, field release and outreach campaign passes its approval gate, delivered with a full SOP manual the client's team operates independently.

voss_orchestrator · system pipeline log
[intake]       Webform submitted → Engagement ID: CC-2026-0142
[apps-script]  Provisioning Drive tree: 03 Engagements/CC-2026-0142/ ... OK
[apps-script]  Workbook row created · Tier 2 module form attached
[zapier]       Bridge A8  fired: Google Forms → ClickUp task           (200 OK, 0.8s)
[zapier]       Bridge A16 fired: Google Forms → Drive engagement folder (200 OK)
[zapier]       Bridge A17 fired: Google Forms → Dropbox client folder   (200 OK)
[zapier]       Bridge A18 fired: Google Forms → Engagement Register row (200 OK)
[clickup]      Task created: "CC-2026-0142 — Intake Received"
[apps-script]  Intake validation: contact ✓  location ✓  expected-location-count = 3 ✓
[clickup]      Stage → SCOPE DEVELOPMENT
[zapier]       Bridge A15 fired: status change → Gate Log row (Evidence confirmed: —  Approved by: —)
[gate-keeper]  Awaiting owner approval — Release Gate 01 (proposal template locked)
[gate-keeper]  Approved 10:42 → Proposal issued from approved template (validity 15 days)
[smartlead]    Campaign HTX Houston: 50 sends/day · 0 bounces · suppression check PASS
[prospeo]      124 verified leads → HubSpot contacts synced
[apps-script]  fieldReleaseCheck(): 3/3 site contacts on file → FIELD RELEASE: PASS
[audit-log]    State synced across 6 bridges in 1.4s · run #418 logged
ClickUp · Engagement Board · Gate Log
A B C D
Stage What must be true before it leaves this stage Automation Status
New Opportunity A real prospect, a named contact and an Engagement ID assigned ClickUp auto-ID Done
Intake Sent / Received Intake form back, plus the tier module form for the chosen tier Zapier A8 · A16 · A17 · A18 Done
Scope Development Scope written, locations known, expected location count recorded Apps Script validation Done
Awaiting Owner Approval Owner has everything needed in one place, owner only Release Gate 01 Approved
Proposal Sent / Awaiting Acceptance Proposal issued from the approved template; validity 15 days Template engine In progress
Agreement & Deposit Signed agreement and 50% deposit cleared (Tier 3: first month invoiced) HubSpot deal stage Pending
Field Planning May proceed outside the gate while site contacts are still being collected Site Contact form Pending
Field Release Pending The gate: complete site-contact set + fieldReleaseCheck PASS Apps Script · Zapier A15 Blocked until gate clears
Technical Mastery

Core Capabilities

A deep dive into the engineering practices that keep automated systems and scrapers running long-term: logging, retries, approval gates and monitoring.

📊

Data Extraction

Expert structuring of complex pages: extracts data from nested tables, multi-layered dashboards, infinite scrolls, lazy loading components, and pages relying entirely on heavy client-side AJAX/JSON rendering.

⚙️

Automation

Fully emulate human browser activity: automated logins, page navigation sequences, scheduled form fills, multi-step checkout clicks, report generations, and time-delayed cron automations.

📥

Bulk Downloads

Automate large-scale assets collection: download files (PDFs, high-resolution imagery, video streams, source documents) concurrently, package files into ZIP structures, and generate structured directories.

🔌

Export Options

Receive your data in standard, clean formats. We specialize in outputting custom Excel sheets (multi-tabbed and color-coded), raw CSV files, optimized databases, or JSON structures ready to feed directly into your APIs.

🔍

Reverse Engineering

Bypass heavy browser loads. By inspecting network exchanges, analyzing WebSockets, and tracking authentication tokens, we build scrapers that fetch data directly from hidden target APIs, reducing bandwidth and failure rates.

💻

Custom Development

Need automation tailored to a niche platform? From proprietary ERP systems to obscure web applications, we analyze, trace, and construct specialized headless automation scripts that align with your business logic.

Ready Databases

Premium Leads Marketplace

We offer pre-packaged historical databases from major platforms, as well as on-demand real-time scraping. Custom scraping systems cost $50 – $2,000+ and leads pricing starts from $0.1 – $0.5+ per lead depending on complexity.
🔥 Huge discounts available if you order any of the featured sample databases below! For new custom scrapers or new target sites, charges will be decided later after feasibility checks.

Creators

Target Volume Range
Emails Verified 94%
Delivery Options Fresh & Historical

Instagram, TikTok, and YouTube creators with pricing tiers, follower counts, verified emails, and niche categorizations.

Creators

Target Volume Range
Emails Verified 92%
Delivery Options Fresh & Historical

Target creator profile details, engagement index parameters, follower metrics, and direct outreach email addresses.

Creators

Target Volume Range
Emails Verified 95%
Delivery Options Fresh & Historical

Social media creators containing bios, category tags, follower stats, channel indicators, and verified emails.

E-Commerce

Target Volume Range
Data Enriched 100%
Contact Method Via Store Page Link

Top Amazon merchants with store names, seller ratings, product listings, category divisions, and direct store page links.

E-Commerce

Target Volume Range
Data Enriched 100%
Contact Method Via Store Page Link

Active eBay merchant store links, seller usernames, feedback percentages, physical locations, and platform contact URLs.

E-Commerce

Target Volume Range
Data Enriched 100%
Contact Method Via Store Page Link

AliExpress store names, store IDs, merchant ratings, main product divisions, and direct supplier store page links.

E-Commerce

Target Volume Range
Data Enriched 100%
Contact Method Via Store Page Link

Alibaba manufacturers and export factories, business types, city headquarters, certifications, and supplier contact page links.

E-Commerce

Target Volume Range
Data Enriched 100%
Contact Method Via Store Page Link

Etsy shop owners, product niche divisions, total sales counts, locations, shop titles, and shop contact links.

E-Commerce

Target Volume Range
Data Enriched 100%
Contact Method Via Store Page Link

Walmart third-party merchant links, business owners, ratings, listing sizes, categories, and direct seller links.

E-Commerce

Target Volume Range
Data Enriched 100%
Contact Method Via Store Page Link

Fast-growing Temu manufacturing merchants, company names, store categories, reviews, and direct merchant contact pages.

E-Commerce

Target Volume Range
Data Enriched 100%
Contact Method Via Store Page Link

DH Gate wholesale factories and trade partners, merchant tiers, positive feedback scores, and supplier profile links.

E-Commerce

Target Volume Range
Data Enriched 100%
Contact Method Via Store Page Link

South Asian merchants listing on Daraz, store ratings, physical cities, follower counts, and direct shop links.

E-Commerce

Target Volume Range
Data Enriched 100%
Contact Method Via Store Page Link

Industrial manufacturers and global suppliers from Made-in-China with verified company store pages and directory links.

Client Reviews

What Our Clients Say

Read reviews from businesses and marketing teams who rely on our custom scrapers and verified lead databases to power their growth.

"ApexAutomation built a custom scraper for a heavy JS-rendered site that other developers said was impossible. Absolute game changer for our lead generation pipeline. Flawless execution and extremely fast delivery."

👨‍💻

Alex Rivera

CTO, LeadScale Media

"We've been purchasing verified Shopify store directories and active e-commerce merchant databases from their Leads Marketplace. The categorization and data accuracy are outstanding. Highly recommend!"

👩‍💻

Jessica Miller

Founder, EcomPulse

"Their headless browser automation tools saved our operations team over 30 hours of manual data extraction every single week. The bypass mechanisms for captchas are incredibly resilient. Outstanding service."

👨‍💼

David K.

Growth Director, FinTech Solutions

"The automated retail scraper they deployed has been flawless. It parses daily price changes across thousands of products and alerts our catalog managers. Highly recommend for any retail business."

👩‍💼

Sarah Jenkins

Operations Head, RetailFlow

"Incredibly resilient scraping solutions. Our scraping tasks require bypassing complex login gates and heavy anti-bot security. ApexAutomation handled it easily. Exceptional customer support!"

👨‍💻

Marcus Brody

Tech Lead, DataSphere

"We bought the custom database of US creators and the response rate on our outreach campaigns instantly doubled. Clean, fully-verified lists that are updated constantly. Will buy again!"

👩‍🎨

Elena Rostova

Marketing Lead, OutreachNinja

"Fantastic Shopify and WooCommerce product extractor. It pulled over 50,000 products with variants, images, and prices in under an hour. Saved us weeks of migration effort."

👨‍💼

Daniel Wu

Co-Founder, DropShipify

"Their API extraction service is top-tier. They reverse-engineered a hidden mobile API for us, letting us query clean JSON data directly instead of dealing with brittle HTML parsers."

👩‍💻

Sophia Martinez

Product Manager, RealInsight

"Custom browser automation scripts run exactly as scheduled. They automated our entire LinkedIn and email warm-up sequences, freeing up our sales reps to focus entirely on closing deals."

👨‍💼

Liam Patterson

Founder, LeadGenius

"The scraper built for our supplier directory has run daily for 6 months without a single crash. The auto-retry and email alert systems are robust."

👨‍💻

Lucas Miller

Tech Director, AutomateHQ

"Flawless lead generation databases. They extracted verified merchant contacts from regional directories with 98% accuracy. Outstanding data partner."

👩‍💼

Emily Zhang

VP of Growth, MarketScale

"Outstanding API discovery work. They mapped a complex hidden AJAX endpoint, turning a slow web crawl into a lightning-fast data feed."

👨‍💼

Carlos Mendez

Founder, ShipQuick

"Their automated headless browser workflows save our catalog team 25 hours a week. Extremely professional team and highly resilient code."

👩‍💼

Chloe Laurent

Operations Manager, LuxeBrands

"Excellent custom scrapers. They bypassed complex security filters and Cloudflare walls that blocked all our in-house attempts. Very impressed."

👨‍💻

Ryan Cooper

Product Manager, RealVal

"The creator database we purchased has been the foundation of our outreach success. We got verified contacts, socials, and engagement rates in one clean sheet."

👩‍💻

Nina Patel

Co-Founder, BuzzCreators
Interactive Estimator

Configure & Book Custom Scrapers

Custom scrapers start from $50 to $2,000+ depending on website qualities and project scope. For the cheapest alternative (rates starting from $0.1 – $0.5+ per lead depending on workload), browse our pre-scraped Lead Store.

Estimated Quote

Contact for Quote Estimated delivery: 1-3 Weeks or more

Configuration Summary

  • Base Engine Scraping Included
  • Complexity: Standard website Included
  • Data Output: CSV / Excel file Included
No payment required yet. We will review and contact you with a final proposal.
💡 Special discounts available for needy individuals, non-profits, or small businesses starting out.
Direct Access

Need automation for a specific website?

Send us the website URL along with your requirements, and we will build a custom browser automation or web scraper tailored to your workflow.

Our Feasibility Promise: For any custom website automation request, our team will get in touch with you, perform a feasibility test on your target site, and report back on the exact extent of automation possible.

What to send:

  • 🔗 Website URL
  • What you need automated or extracted
  • 📊 Expected output format (Excel, JSON, etc.)

Email Contact

Copy our address or click below to launch your email client.

Send Email Proposal Response time: as soon as possible
About Us

About Apex Automation Team

Who we are, what we build, and where to find us officially.

Apex Automation Team

Apex Automation Team (apexautomationteam.com) is an independent business automation and web data agency founded by Muhammad Osama. We automate complete business processes: client onboarding and engagement, portal and dashboard operations, document and OCR pipelines, verification, monitoring, reporting and CRM sync. We also build the data layer behind them, including custom web scrapers, browser automation bots, software testing and QA (a trained manual testing team or an automated testing system you own), bulk downloaders, API extraction tools and AI agents, plus ready-made verified lead databases for marketplaces such as Amazon, eBay, Etsy, Alibaba and AliExpress.

Built for your target site Every solution is developed for the client's exact target website and business requirements, with a free feasibility check before any work begins. Founder: Muhammad Osama. Contact: contact@apexautomationteam.com.

Services Worldwide

  • Business Process Automation (workflows, portals, documents, verification, monitoring, reporting)
  • Software Testing & QA (a trained manual testing team, an automated testing system, or both)
  • Custom Web Scraping & Data Extraction
  • Browser Automation Bots
  • Bulk Download Tools
  • API Extraction & Reverse Engineering
  • Custom AI Agents
  • Verified Lead Databases (Amazon, eBay, Etsy, Alibaba, AliExpress & more)

Official Profiles

Projects by Apex Automation Team

  • Chess Arena lets you play chess online against 16 bots, up to Stockfish 18, with full game review.
  • Infinite Cosmos is an interactive 3D journey from the Big Bang to the present universe.

Read the full FAQ →  ·  Full About page →  ·  Guides →

What is Apex Automation Team?

Apex Automation Team (apexautomationteam.com) is an independent business automation and web data agency founded by Muhammad Osama. We automate complete business processes: client onboarding, portal and dashboard operations, document and OCR pipelines, verification, monitoring and reporting. We also build the data layer behind them, including custom web scrapers, browser automation bots, software testing and QA (a trained manual testing team or an automated testing system you own), bulk downloaders, API extraction tools, AI agents and ready-made verified lead databases. Services are delivered to clients worldwide.

What services does Apex Automation Team offer?

Business process automation (workflow orchestration, portal and dashboard operations, document and PDF/OCR pipelines, verification, monitoring, reporting and CRM sync), custom web scraping, browser automation, software testing and QA with a trained manual team or an automated suite, bulk download tools, API extraction and reverse engineering, custom AI agents, and verified lead databases for Amazon, eBay, Etsy, Alibaba, AliExpress, Walmart, Temu, Daraz sellers and influencers. Every tool is built for your exact process and target website.

Do you automate whole business processes or just single tasks?

Whole processes. A typical engagement replaces an entire process with orchestrated bots and AI agents: intake, the work itself, the checks, the records and the reporting. We leave an approval step wherever a decision is irreversible. Samples from our own work include an Enterprise Operations OS for client engagement, reseller portal automation, PDF/OCR document pipelines, automated VIN verification with sheet matching and a live Google Sheets monitor with leak alerts.

Do you test software as well as automate it?

Yes, and there are two ways to take it. Option 1 is a trained manual testing team that runs cycles on demand: functional, exploratory and regression testing, cross-browser and cross-device checks, checkout and payment flows, accessibility and localisation, with every bug reported with steps, video, console detail and severity. Option 2 is an automated testing system built for you and owned by you, covering the stable flows on every release, wired into your CI and reported into Jira, ClickUp or Slack. Take either one, or both together.

Is Apex Automation Team related to other companies called “Apex Automation”?

No. Apex Automation Team is an independent web automation and data agency operating only at apexautomationteam.com. It is not affiliated with industrial automation, electrical engineering or Oracle APEX products that share a similar name.

How do I order a custom scraper or automation?

Email contact@apexautomationteam.com with the target website URL, what you need automated or extracted, and your preferred output format (Excel, CSV, JSON, Google Sheets). We run a free feasibility test on the target site and reply with a quote as soon as possible.

Which websites can Apex Automation Team automate?

Almost any website, including JavaScript-heavy, login-protected, infinite-scroll and Cloudflare-protected targets. Our team tests every target site first and reports exactly how much can be automated before you pay.

Where can I follow Apex Automation Team online?

Official profiles: YouTube @apexautomationteam, X @apexautomation3, TikTok @apexautomationteam, Instagram @apexautomation2026, Facebook (Apex Automation Team) and the founder's LinkedIn profile. Side projects built by the team include Chess Arena (chess.apexautomationteam.com) and Infinite Cosmos (bigbang.apexautomationteam.com).