Senior Data & Python Software Engineer

Ceartas
Berlin, Germany
13 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
2 years minimum
Working hours
Regular working hours
Job source

Tech stack

Adaptable Database Systems Application Programming Interfaces (APIs) Airflow Amazon Web Services Microsoft Azure Cloud Computing Data as a Services Web Scraping Data Transformation Data Mining Relational Databases Database Queries
+18 more
Database Schema Software Debugging Django Web Framework Python (Programming Language) PostgreSQL Performance Tuning Selenium Software Deployment Data Logging Reliability of Systems Backend Fastapi Containerization Information Technology Playwright Restful APIs Data Pipelines Docker

Job description

brand protection workflows. Ensure that data moves reliably from collection through processing to storage while maintaining performance, resilience, and operational stability at scale.

Ensure Data Quality and Governance:

Own data validation, consistency, and governance across ingestion, storage, and serving layers.

Establish clear standards for schema design, transformation logic, and monitoring to guarantee

trustworthy, production-grade datasets that can be reliably consumed across the organization.

Optimize Performance and Reliability:

Continuously improve scraping system efficiency through performance tuning, cost optimization, and architectural enhancements. Implement logging, metrics, and tracing to monitor production systems, diagnose issues quickly, and maintain high reliability under growing workloads., * Design, build, and maintain high-performance web scraping systems as well backend services and data pipelines supporting web data extraction and brand protection use cases

  • Implement and maintain scraping focused APIs and other data services that power internal products and external integrations
  • Build reliable ingestion, processing, and storage workflows for large-scale web data
  • Handle cleaning of web data and ensure data quality, validation, and governance across ingestion, storage, and serving layers
  • Optimize scraping systems for performance, scalability, reliability, and cost efficiency
  • Monitor, debug, and improve scraping system reliability using observability tools (logging, metrics, tracing)
  • Collaborate closely with product and engineering teams to deliver features from design through full end-to-end production deployment
  • Take independent ownership of systems in production, including maintenance,
  • iteration and performance management

Requirements

  • Experience with web scraping
  • Strong SQL skills
  • Strong Python experience
  • Experience with PostgreSQL or similar relational databases
  • Experience designing and building scalable APIs and backend services (e.g. FastAPI, Django, or similar frameworks)
  • Experience designing efficient, scalable data models and database schemas
  • Hands-on experience deploying and operating systems in the cloud (AWS, GCP, or Azure)
  • Experience working with Docker and containerized environments Preferred Technical Requirements:

  • Experience with workflow orchestration tools such as Airflow
  • Experience with browser-based automation tools (Playwright, Selenium, or similar)
  • Experience with DBT or analytics-focused data transformation workflows
  • Experience building or operating high-concurrency systems and task queues
  • Experience designing and deploying cloud-native workflows on AWS
  • Familiarity with CI/CD pipelines and production deployment practices
  • Experience working in a high-growth, early-stage startup environment
  • Experience - University education in a technical field such as Computer Science, Engineering or similar. Masters level preferred. 4+ years ( or 2 year+ in a early stage startup)

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on de.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

3:33 min

Connecting frontends via a FastAPI proxy backend layer

Saoussen Chaabnia Saoussen Chaabnia · Europe 2026 Virtual

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

1:42 min

Introduction to the fast API web framework

Sebastián Ramírez · World Congress 2022

Videos

See all

Related articles

See all