Web Scraper

Nielseniq
Barcelona, Spain
6 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
3 years minimum
Working hours
Regular working hours

Tech stack

HTML JavaScript (Programming Language) Amazon Web Services Microsoft Azure Bash Shell Code Review Data Validation Linux Django Web Framework Python (Programming Language) Regular Expressions Software Engineering
+11 more
Web Application Frameworks Time Series Databases Git Fastapi Pandas Information Technology Influxdb Graphql Restful APIs Data Pipelines Docker

Job description

Experteer Overview In this role you will help optimize data collection as a Python Developer in our scraping team.You will design, implement, and document robust Scrapy spiders to withstand website changes while preserving existing functionality.You’ll ensure high data quality through code reviews and validation, and build sophisticated crawlers that navigate anti-bot measures.You will contribute to major developments across codebases and share knowledge through documentation and training.This is a chance to shape scalable data pipelines for real-time market visibility in a fast-growing, startup-minded environment.Compensaciones / Beneficios* Design, implement and document Scrapy spiders for robust data collection* Perform code reviews and data validation to ensure quality* Develop crawlers that handle anti-bot measures using HTTP and browser techniques* Architect and contribute to new developments across multiple codebases with documentation and training* Collaborate with cross-functional teams to share knowledge and enable tooling improvementsResponsabilidades* Master in Computer Science, IT, or related field* At least 3 years of professional software engineering experience* Knowledge of BeautifulSoup or Scrapy* Familiarity with HTML and JavaScript, understanding of SPAs* Experience with RESTful and/or GraphQL APIs* Hands-on experience with Django, FastAPI or similar Python frameworks* Scrapy knowledge is a plus* Experience with time series databases like InfluxDB is advantageous* Strong skills in Docker, Git, pandas, regex, Linux, and bash scripting* Proven experience with AWS, GCP, or AzureRequisitos principales* Flexible working environment* Volunteer time off* LinkedIn Learning* Employee-Assistance-Program (EAP)

Requirements

  • Master in Computer Science, IT, or related field
  • At least 3 years of professional software engineering experience
  • Knowledge of BeautifulSoup or Scrapy
  • Familiarity with HTML and JavaScript, understanding of SPAs
  • Experience with RESTful and/or GraphQL APIs
  • Hands-on experience with Django, FastAPI or similar Python frameworks
  • Scrapy knowledge is a plus
  • Experience with time series databases like InfluxDB is advantageous
  • Strong skills in Docker, Git, pandas, regex, Linux, and bash scripting
  • Proven experience with AWS, GCP, or Azure Requisitos principales

  • Flexible working environment
  • Volunteer time off
  • LinkedIn Learning
  • Employee-Assistance-Program (EAP)

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

5:54 min

The technical evolution of modern web scraping infrastructure

Chris Heilmann +4 · LIVE

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:21 min

Projecting external HTML content using default and named slots

Rowdy Rabouw Rowdy Rabouw · WWC 2022

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

3:02 min

Assembling the tech stack for scalable indexing workflows

Chris Heilmann +2 · LIVE

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

Videos

See all

Related articles

See all