Senior Engineer for Web Automation Project

Job ID: 40271586

Budget: $10,000 – $20,000 USD

Senior Software Engineer (Web Automation & Data Synchronization)

Company: Confidential
Role: Senior Web Automation Engineer
Location: Remote (Global Pod)
Type: Full-Time Contract

1. About the Project

We are developing a high-velocity national procurement and market intelligence platform for the $100B construction finishing industry, and we are aiming to scale our core "Data Factory" to support a 50-state national launch.

The objective is to synchronize multi-variate inventory, pricing, and LTL logistics data from 100+ national retailers into a unified, high-performance data schema.

2. Key Responsibilities

Automation Architecture: Design and deploy high-concurrency automation clusters using Node.js (TypeScript) and headless browser frameworks (Playwright/Puppeteer).

Stateful Flow Simulation: Develop sophisticated scripts to simulate complex user journeys, including geo-localized session management and extraction of landed costs (Price + Tax + Freight).

Neural Data Normalization: Architect ETL pipelines to transform unstructured web data into our Internal Universal Material Schema using advanced regex and LLM-assisted parsing.

Perceptual Indexing: Implement pHash (Perceptual Hashing) logic to identify and de-duplicate identical physical products across the national market.

Infrastructure Management: Manage auto-scaling containerized workloads on AWS (Fargate/ECS) with advanced proxy mesh orchestration.

3. Technical Requirements

Core: Expert-level Node.js (TypeScript) or Python.

Automation: Mastery of Playwright or Puppeteer.

System Resilience: Deep understanding of request orchestration, session persistence, and residential proxy management to ensure high uptime against enterprise-grade site architectures.

Cloud: AWS (Fargate, Lambda, S3).

Vision: Familiarity with image processing/fingerprinting (OpenCV/pHash) is a significant advantage.

4. The Challenge: National Scale

You will be building a system designed to support a $10MM+ ARR run-rate within 12 months. This requires a "Zero-Error" approach to data integrity, specifically ensuring that material quantities and freight weights remain accurate across 40,000+ US zip codes.

5. Why Join Us?

Lead the technical build of a proprietary data moat for a venture-backed project.

Tackle high-complexity architectural challenges involving distributed state management.

Milestone-based compensation with a clear path to long-term leadership as the team scales.

How to Apply

Please provide your GitHub profile or a portfolio highlighting your work in browser automation. Specifically, describe a project where you managed high-concurrency data synchronization across multiple state-dependent web environments.
Related categories: Data Processing Cloud Computing Node.js