Senior Web Scraper | Beautifulsoup Or Scrapy

Nielseniq
Barcelona, Spain
4 days ago
Apply on www.buscojobs.com.es
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
3 years minimum
Working hours
Regular working hours

Tech stack

HTML JavaScript (Programming Language) Amazon Web Services Microsoft Azure Bash Shell Code Review Data Validation Linux Django Web Framework Python (Programming Language) Regular Expressions Software Engineering
+12 more
Web Application Frameworks Web Crawlers Time Series Databases Git Fastapi Pandas Single Page Application Information Technology Influxdb Graphql Restful APIs Docker

Job description

Experteer Overview As Senior Python Developer in our scraping team, you will optimize data collection by building robust Scrapy spiders and ensuring resilient crawlers against website changes.You will uphold code and data quality through reviews and validation, and design sophisticated crawling solutions that navigate anti-bot measures.You’ll contribute to major builds across codebases, documenting clearly and sharing knowledge with relevant teams.This role sits in a fast-growing, start-up-spirited NIQ for Digital Commerce, focused on real-time retail insights.Compensaciones / Beneficios * Design, implement and document robust Scrapy spiders * Perform code reviews and data validation to ensure quality * Develop web crawling solutions that bypass advanced anti-bot measures * Architect new features across multiple codebases with clear documentation * Provide training and knowledge sharing to relevant teams Responsabilidades * Master’s degree in Computer Science, IT, or related field * At least 3 years of professional software engineering experience * Knowledge of BeautifulSoup or Scrapy * Familiarity with HTML and JavaScript, understanding of single-page applications * Experience with RESTful and/or GraphQL APIs * Hands-on experience with Django, FastAPI, or similar Python web frameworks * Knowledge of Scrapy is a plus * Experience with time series databases (InfluxDB) is advantageous * Strong skills in Docker, Git, pandas, regular expressions, Linux, and bash scripting * Experience with AWS, GCP, or Azure * Enthusiastic, motivated, autonomous; enjoys tackling challenging software problems * Passionate about tech and continuous learning Requisitos principales * Flexible working environment * Volunteer time off * LinkedIn Learning * Employee-Assistance-Program (EAP)

Requirements

Compensaciones / Beneficios * Design, implement and document robust Scrapy spiders * Perform code reviews and data validation to ensure quality * Develop web crawling solutions that bypass advanced anti-bot measures * Architect new features across multiple codebases with clear documentation * Provide training and knowledge sharing to relevant teams Responsabilidades * Master’s degree in Computer Science, IT, or related field * At least 3 years of professional software engineering experience * Knowledge of BeautifulSoup or Scrapy * Familiarity with HTML and JavaScript, understanding of single-page applications * Experience with RESTful and/or GraphQL APIs * Hands-on experience with Django, FastAPI, or similar Python web frameworks * Knowledge of Scrapy is a plus * Experience with time series databases (InfluxDB) is advantageous * Strong skills in Docker, Git, pandas, regular expressions, Linux, and bash scripting * Experience with AWS, GCP, or Azure * Enthusiastic, motivated, autonomous; enjoys tackling challenging software problems * Passionate about tech and continuous learning Requisitos principales * Flexible working environment * Volunteer time off * LinkedIn Learning * Employee-Assistance-Program (EAP)

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

5:54 min

The technical evolution of modern web scraping infrastructure

Chris Heilmann +4 · LIVE

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:21 min

Projecting external HTML content using default and named slots

Rowdy Rabouw Rowdy Rabouw · World Congress 2022

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

3:43 min

Scaling web scraping infrastructure to bypass strict security restrictions

Tim Ruscica · Coffee With Developers

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

Videos

See all

Related articles

See all