Web Scraping Engineer

ALPHAMATICIAN LLC
United States
3 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
2 years minimum
Compensation
$120,000.0 - $140,000.0
Working hours
Shift work
Job source

Tech stack

PHP (Programming Language) CodeIgniter Web Scraping Python (Programming Language) MySQL Node.Js Selenium Captcha Puppeteer (Software) Playwright

Job description

  • Morning (7am ET). Review overnight pipeline logs. Identify failures, anomalies, or coverage gaps. Triage fixes and follow-ups.
  • Engineering work. Fix PHP bugs. Update scrapers as target sites change their structure or defenses. Update cron logic. Ship incremental improvements to collection coverage and quality.
  • Data quality. Run checks against current and historical baselines to confirm coverage and accuracy.
  • Client questions. Respond to client inquiries about coverage, methodology, or anomalies as they come in.
  • Strategy. Periodically discuss collection strategies, help scope and stand up new datasets, and contribute to new products and features.

This is operational work with a steady rhythm. The reward is in keeping an important data product running well, and in being good at a specific kind of hard problem (scraping hard sites at scale) that few people are good at., * Production PHP experience. CodeIgniter 4 is a strong plus. You must be able to point to public PHP work: a repo, contributions to a project, a blog post, or similar.

Requirements

Do you have experience in Python?, * Python proficiency with modern scraping libraries. Working fluency in Playwright, Scrapy, Selenium, Requests, httpx, BeautifulSoup, or comparable. Real scraping work lives in this toolkit.

  • Demonstrated experience scraping hard targets at scale. Sites with active anti-bot defenses, dynamic rendering, rate-limit walls, or aggressive blocking. You must include a link to public scraping work in your application, or describe a specific scraper you built in detail (target, defenses encountered, how you solved them).
  • MySQL competence. Reading and writing non-trivial queries against tables with hundreds of millions of rows.
  • Schedule. 7am US Eastern start.
  • US-based with verifiable employment history.

Strong plusses

  • Direct experience with anti-bot evasion: residential proxies, TLS fingerprint matching, JA3, header rotation, CAPTCHA strategy.
  • Comfort with mature, incrementally maintained codebases.
  • Background in financial data, alternative data, or equity research.
  • Node.js, Puppeteer, or additional automation tooling., * production PHP: 3 years (Required)
  • production web scraping: 2 years (Required)

Benefits & conditions

$120,000 - $140,000 a year - Full-time, Contract, Pulled from the full job description

  • Flexible schedule, * Flexible schedule

About the company

Alphamatician is hiring a Web Scraping Engineer for an individual-contributor role at the heart of our data operations. The work is technical, operational, and quietly important: keeping a mature alternative-data product running reliably for institutional investors and decision makers. The role is structured as contract-to-hire, with W-2 conversion after a successful initial engagement.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:40 min

Overcoming modern anti-bot mechanisms and network access restrictions

Vidas Bacevičius Vidas Bacevičius · WWC 2025

3:43 min

Scaling web scraping infrastructure to bypass strict security restrictions

Tim Ruscica · Coffee With Developers

45 sec

Working securely with Node.js path application programming interfaces

Sonya Moisset · WWC 2023

2:18 min

Scaling MySQL databases for massive user growth

Johannes Nicolai Johannes Nicolai +1 · LIVE

2:01 min

Bypassing anti-bot protections with proxies and emulation

Jan Curn Jan Curn · WWC 2025

5:54 min

The technical evolution of modern web scraping infrastructure

Chris Heilmann +4 · LIVE

Videos

See all

Related articles

See all