Business Development

Alphamatician

Oregon (WI)

On-site

USD 100,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Alphamatician is seeking a Web Scraping Engineer to own day-to-day data collection and quality for an institutional data product. This individual-contributor role is embedded in our data operations and focuses on reliable scraping at scale, handling hard targets and evolving sites.

You will review overnight pipelines, fix scraping logic, run quality checks, and respond to client inquiries. The role emphasizes practical problem solving on a mature data platform with continual improvements.

Qualifications

  • Production PHP experience with public work to cite (repo/blog/post).
  • Python proficiency with modern scraping libraries (Playwright, Scrapy, Selenium, Requests, httpx, BeautifulSoup).
  • Experience scraping hard targets at scale with anti-bot defenses and dynamic content; must cite a scraper built in detail.
  • MySQL reading/writing and optimizing complex queries on large datasets.
  • US-based with verifiable employment history.
  • Starts at 7am US Eastern.

Responsibilities

  • Review overnight pipeline logs and triage failures or gaps.
  • Fix PHP bugs and update scrapers as sites change structure or defenses.
  • Update cron logic and ship incremental improvements to collection coverage and quality.
  • Run data quality checks against current/historical baselines for coverage and accuracy.
  • Answer client inquiries about coverage, methodology, or anomalies.
  • Discuss collection strategies and contribute to new datasets and product features.

Skills

Working fluency in English
Attention to detail

Tools

PHP (CodeIgniter 4)
Playwright
Scrapy
Selenium
Requests
httpx
BeautifulSoup
MySQL

Job description

Alphamatician is hiring a Web Scraping Engineer for an individual-contributor role at the heart of our data operations. The work is technical, operational, and quietly important: keeping a mature alternative-data product running reliably for institutional investors.

What You Will Do
  • Morning (7am ET). Review overnight pipeline logs. Identify failures, anomalies, or coverage gaps. Triage fixes and follow-ups.
  • Engineering work. Fix PHP bugs. Update scrapers as target sites change their structure or defenses. Update cron logic. Ship incremental improvements to collection coverage and quality.
  • Data quality. Run checks against current and historical baselines to confirm coverage and accuracy.
  • Client questions. Respond to client inquiries about coverage, methodology, or anomalies as they come in.
  • Strategy. Periodically discuss collection strategies, help scope and stand up new datasets, and contribute to new products and features.

This is operational work with a steady rhythm. The reward is in keeping an important data product running well, and in being good at a specific kind of hard problem (scraping hard sites at scale) that few people are actually good at.

Hard Requirements
  • Production PHP experience. CodeIgniter 4 is a strong plus. You must be able to point to public PHP work: a repo, contributions to a project, a blog post, or similar.
  • Python proficiency with modern scraping libraries. Working fluency in Playwright, Scrapy, Selenium, Requests, httpx, BeautifulSoup, or comparable.
  • Demonstrated experience scraping hard targets at scale. Sites with active anti-bot defenses, dynamic rendering, rate-limit walls, or aggressive blocking. You must include a link to public scraping work in your application, or describe a specific scraper you built in detail (target, defenses encountered, how you solved them).
  • MySQL competence. Reading, writing, and optimizing non-trivial queries against tables with hundreds of millions of rows.
  • Schedule. 7am US Eastern start.
  • US-based with verifiable employment history.
Strong Plusses
  • Direct experience with anti-bot evasion: residential proxies, TLS fingerprint matching, JA3, header rotation, CAPTCHA strategy.
  • Comfort with mature, incrementally maintained codebases rather than green-field environments.
  • Background in financial data, alternative data, or equity research.
  • Node.js, Puppeteer, or additional automation tooling.
Scope

This role focuses on day-to-day data collection, quality, and client-facing operations. It is not an architecture, modernization, or infrastructure-ownership role. If you want to redesign systems or own the stack, this is not the right fit. If you want to do excellent scraping work, solve real bugs, and own the daily craft of keeping a data product reliable.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Web Scraping Engineer
Web Scraping Engineer

Alphamatician • Indiana (PA)

On-site
USD 120,000 - 180,000
Data Analyst
Data Analyst

Alphamatician • Indiana (PA)

On-site
USD 95,000 - 150,000
Web Scraping Engineer & Data Quality Specialist
Web Scraping Engineer & Data Quality Specialist

Alphamatician • Indiana (PA)

On-site
USD 95,000 - 150,000
Web Scraping Engineer – Data Reliability & Ops
Web Scraping Engineer – Data Reliability & Ops

Alphamatician • Indiana (PA)

On-site
USD 120,000 - 180,000
Web Scraping Engineer — Data Quality & Reliable Pipelines
Web Scraping Engineer — Data Quality & Reliable Pipelines

Alphamatician • Oregon (WI)

On-site
USD 100,000 - 150,000
Senior Data Engineer - Web Scraping
Senior Data Engineer - Web Scraping

Jobgether • United States

Remote
USD 120,000 - 170,000
Fully remote
Full-time
Autonomy
+1
iOS/Web App Developer
iOS/Web App Developer

Worky • United States

On-site
USD 165,000 - 220,000
Senior Data Engineer - Web Scraping
Senior Data Engineer - Web Scraping

Triwill Group • Spain (TX)

On-site
USD 75,000 - 110,000
Fully remote
Full-time
Autonomy and ownership
+1
Director of Software Engineering (Node.js & Web Scraping Expert)
Director of Software Engineering (Node.js & Web Scraping Expert)

Roman Health Pharmacy LLC • Los Angeles (CA)

On-site
USD 150,000 - 200,000
Health insurance
Flexible working hours
Professional development opportunities
Senior Data Engineer
Senior Data Engineer

Rachel Paul Recruiting • New York (NY)

Hybrid
USD 120,000 - 150,000