Senior Software Engineer (Web Scraping, Data Acquisition, Data Collection)

Nielseniq
Barcelona, Spain
1 day ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
5 years minimum
Working hours
Regular working hours

Tech stack

JavaScript (Programming Language) Application Programming Interfaces (APIs) Databases Data Centers Web Scraping Software Debugging Python (Programming Language) MongoDB Redis Reverse Engineering Selenium Session Management
+7 more
Parquet Build Management Influxdb Playwright Avro Mitmproxy Stream Processing

Job description

Experteer Overview La información a continuación detalla los requisitos del puesto, la experiencia esperada del candidato y las cualificaciones correspondientes.In this role you will design and operate scalable distributed scraping systems to extract high-quality data from thousands of domains.You will work within the Scraping and Ru**amp;D team to push innovative approaches and share knowledge, collaborating with Customer Success, Operations, and Data.You’ll tackle anti-bot challenges and build monitoring to prevent blockers, contributing to fast, scalable visibility for retailers and brands.This is a hands-on, problem?solving role for someone who loves turning difficult data problems into reliable solutions.Compensaciones / Beneficios* Design and build scalable distributed scraping architectures for thousands of domains* Reverse engineer APIs and mobile apps using tools like Frida, mitmproxy, Charles Proxy, burp* Detect and bypass anti-bot measures to create human-like crawlers* Lead technical innovation in scraping, prototype approaches, share knowledge with the team* Set up anomaly detection and monitoring to catch blockers or failures earlyResponsabilidades* 5+ years of Python experience and involvement in large-scale scraping projects* Deep understanding of anti-bot technologies and bypass techniques, fingerprinting, WAFs, JavaScript challenges* Experience with Scrapy, Requests, httpx, Selenium, Playwright or similar headless drivers* Proficiency with proxies (residential, datacenter, rotating), session management, cookie injection, xqbhyrx TLS tweaking* Strong debugging skills across browser, HTTP traffic, and device-level interactions* Independent problem solver with a track record of solving new problems* Familiarity with Parquet, Avro or real-time stream processing* Knowledge of databases like MongoDB, InfluxDB, and Redis and how to optimize themRequisitos principales* Flexible working environment* Volunteer time off* LinkedIn Learning* Employee-Assistance-Program (EAP)

Requirements

This is a hands-on, problem?solving role for someone who loves turning difficult data problems into reliable solutions.Compensaciones / Beneficios* Design and build scalable distributed scraping architectures for thousands of domains* Reverse engineer APIs and mobile apps using tools like Frida, mitmproxy, Charles Proxy, burp* Detect and bypass anti-bot measures to create human-like crawlers* Lead technical innovation in scraping, prototype approaches, share knowledge with the team* Set up anomaly detection and monitoring to catch blockers or failures earlyResponsabilidades* 5+ years of Python experience and involvement in large-scale scraping projects* Deep understanding of anti-bot technologies and bypass techniques, fingerprinting, WAFs, JavaScript challenges* Experience with Scrapy, Requests, httpx, Selenium, Playwright or similar headless drivers* Proficiency with proxies (residential, datacenter, rotating), session management, cookie injection, xqbhyrx TLS tweaking* Strong debugging skills across browser, HTTP traffic, and device-level interactions* Independent problem solver with a track record of solving new problems* Familiarity with Parquet, Avro or real-time stream processing* Knowledge of databases like MongoDB, InfluxDB, and Redis and how to optimize themRequisitos principales* Flexible working environment* Volunteer time off* LinkedIn Learning* Employee-Assistance-Program (EAP)

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

5:54 min

The technical evolution of modern web scraping infrastructure

Chris Heilmann +4 Ā· LIVE

2:01 min

Migrating existing applications from MongoDB to Postgres

Nikita Shamgunov Nikita Shamgunov Ā· WWC 2024

3:55 min

Demonstrating semantic routing thresholds with the Redis vector library

3:02 min

Audience Q&A on data formats and engine tradeoffs

Matthias Niehoff Matthias Niehoff Ā· WWC Europe 2026

1:21 min

Realizing the limitations of MongoDB for live statistics

Josip Stuhli Josip Stuhli Ā· WWC 2023

3:42 min

Comparing in-memory and Redis storage for cache scalability

Simone Sanfratello Ā· WWC 2022

Videos

See all

Related articles

See all