iOS/Web App Developer

Worky

United States

On-site

USD 165,000 - 220,000

Part time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Worky is seeking a contractor to rebuild a deal-sourcing pipeline. You will build a chain scraper to visit deal pages, extract structured data, and perform two-way synchronization with the database.

The work includes an LLM extraction layer using Claude for structured output and a deterministic core for validation and pricing calculations. The role emphasizes architecture depth, automated testing, and observable pipelines with weekly orchestration via GitHub Actions.

Qualifications

  • Experience building end-to-end data pipelines with two-way synchronization and idempotency.
  • Proficient with Playwright, TypeScript/Node.js, and Postgres ecosystems.
  • Experience designing and testing robust data pipelines with auditability and observability.

Responsibilities

  • Visit chain deal pages, extract structured deals, and sync with the database, ensuring idempotent operations.
  • Fetch page HTML and screenshots; leverage LLMs for structured extraction guided by configuration.
  • Build deterministic core functions for validation, pricing, categorization, and scoring; ensure repeatable outputs.

Skills

Strong architectural mindset
Self-directed
Problem solving
Attention to data quality

Tools

Playwright
TypeScript
Node.js
SQL
PostgreSQL
REST APIs
Supabase
GitHub Actions
LLM API integration
ETL & data-pipeline

Job description

Engagement: Contract, approximately 120–160 hours. Full-time focus preferred.
Location: Remote

The Role

We're rebuilding a deal-sourcing pipeline. The system is currently an opaque third-party AI agent on Tasklet. We need inspectable, version-controlled code.

What You'll Build
Chain scraper

Visit every chain's deal page, extract structured deals, and perform a two-way sync with the database. This includes inserting new deals, updating changed deals, and deactivating stale deals. The process must be idempotent and re-runnable.

LLM extraction layer

Fetch page HTML and screenshots, then use Claude to extract structured deal objects guided by per-chain configuration.

Deterministic core

Build pure, unit-tested TypeScript functions for validation rules, pricing math using dp, op, and pct, category mapping, and 0–100 scoring. The same input must produce the same output every time.

Audit steps

Build link-health checks, content verification to confirm that database records match live pages, data-quality audits for issues such as fake free offers, incorrect categories, math errors, and duplicates, and CTA audits.

Orchestration

Create a weekly scheduled pipeline using GitHub Actions cron that runs every step in order and produces both machine-readable and human-readable run reports.

Observability

Store structured logs in Postgres for every run and every chain so we can always answer questions such as, “What did the pipeline do last Monday, and why?”

Submission enrichment

Build a Supabase Edge Function that turns free-text user submissions into structured and scored deals.

Required Skills
Playwright

This is the primary scraping mechanism and the hardest part of the job. You should be comfortable driving headless browsers against JavaScript-rendered and bot-resistant sites and debugging scrapers when a site changes its markup.

TypeScript and Node.js

The entire pipeline will be a TypeScript and Node.js package. You should write strong, idiomatic TypeScript and be comfortable structuring a production codebase rather than a collection of scripts.

SQL and Postgres

You will design several tables and confidently read, write, compare, and update rows. Supabase Postgres is the backing store.

REST APIs

You should be comfortable consuming and reasoning about REST and PostgREST endpoints.

Automated testing

You should write unit tests for pure functions as a standard part of development, particularly for validation, scoring, and pricing calculations.

Web-scraping judgment

You understand that scrapers are inherently fragile, plan for website changes, and know how to isolate configuration so repairs are inexpensive.

Strongly Preferred
LLM API integration

Experience using Claude or a similar model for structured extraction with deterministic post-processing. You should understand prompt and schema design and be conscious of API costs.

Supabase

Experience with Edge Functions, pg_cron, migrations, and service-role versus anonymous key handling.

CI/CD automation

Experience with GitHub Actions, including scheduled workflows, secrets, and artifacts.

ETL and data-pipeline experience

Experience with two-way synchronization, idempotency, diffing, and data-quality auditing.

Nice to Have

Experience building self-healing or configuration-driven scrapers that can adapt to site changes with minimal code edits.

Experience integrating Slack or webhook notifications for pipeline reporting.

Who This Is Right For

This role is for someone who is genuinely comfortable with architecture and automation depth, not just SQL.

The value is not in writing queries. It is in building a robust, observable, and testable pipeline while taming a fragile scraping layer. If scraping changing websites sounds like a familiar, solvable but fiddly problem rather than a mystery, you may be a strong fit.

You should be self-directed. The brief is detailed, the requirements are clear, and we want someone who can take the project from beginning to end while making sound decisions about open questions such as runtime, extraction engine, and project phasing.

Scope and Timeline
Estimate

Approximately 120–160 hours of focused work. A capable developer using modern AI tooling may be able to establish the basic end-to-end loop in roughly five to seven days. The remaining work will involve making Playwright reliable across approximately 113 sites and ensuring the complete pipeline runs smoothly.

Reality Check on Scraping

The scraper will not be a one-time build. Websites change, and the Playwright scripts will occasionally require adjustments. The initial implementation should make ongoing maintenance as inexpensive as possible through configuration-driven design and detailed logging.

Phase 1

Build the core loop, including configuration, extraction, scraping and two-way synchronization, data-quality auditing, deterministic scoring, run tables and logs, a command-line interface with a dry-run option, and unit tests.

Phase 2

Build link-health, content-verification, and CTA audits; news and DROPs scrapers; submission enrichment; the complete Monday orchestration schedule; and report delivery.

How We'll Evaluate

Demonstrated Playwright experience with real, messy websites through a portfolio, repository, or short technical conversation.

Clean and tested TypeScript. We will want to understand how you structure pure functions and tests.

Sound judgment about scraper fragility, including how you would respond when a restaurant chain changes its page layout or blocks headless browsers.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Business Development
Business Development

Alphamatician • Oregon (WI)

On-site
USD 100,000 - 150,000
Web Scraping Specialist
Web Scraping Specialist

Social Fetch • United States

On-site
USD 120,000 - 170,000
Flexible hours
Unlimited PTO
Equipment stipend
+2
AI Engineer — Research Agents (Full-Stack)
AI Engineer — Research Agents (Full-Stack)

Sixtyfour • San Francisco (CA)

On-site
USD 180,000 - 240,000
Data Analyst
Data Analyst

Alphamatician • Indiana (PA)

On-site
USD 95,000 - 150,000
Director of Software Engineering (Node.js & Web Scraping Expert)
Director of Software Engineering (Node.js & Web Scraping Expert)

Roman Health Pharmacy LLC • Los Angeles (CA)

On-site
USD 150,000 - 200,000
Health insurance
Flexible working hours
Professional development opportunities
Ruby Engineer - Web Scraping (Remote)
Ruby Engineer - Web Scraping (Remote)

SearchApi • Town of Poland (NY)

Remote
USD 100,000 - 130,000
Equity share
Profit sharing
Annual team retreats
Anti-Bot Engineer
Anti-Bot Engineer

SearchApi, LLC • Northern (KY)

Hybrid
USD 120,000 - 180,000
Fully Remote Work
Equity share
Profit sharing
+2
Senior Software Engineer
Senior Software Engineer

Midpage • New York (NY)

On-site
USD 90,000 - 130,000
Back end Software Engineer (web scraping/data acquisition)
Back end Software Engineer (web scraping/data acquisition)

Sapient Search • United States

On-site
USD 120,000 - 180,000
Senior Data Engineer - Web Scraping
Senior Data Engineer - Web Scraping

Jobgether • United States

Remote
USD 120,000 - 170,000
Fully remote
Full-time
Autonomy
+1