> Markdown version of [/jobs/ext/2116797-big-data-engineer](https://www.wearedevelopers.com/jobs/ext/2116797-big-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # big data engineer - **Company:** SimilarWeb LTD - **Location:** United States - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Airflow, Amazon Web Services, Data Analysis, Microsoft Azure, Big Data, Computer Programming, Continuous Integration, Information Engineering, Data Structures, Software Design Patterns, Python (Programming Language), Machine Learning, Object-Oriented Software Development, Parquet, Data Ingestion, Apache Spark, Electronic Medical Records, Git, Containerization, Kubernetes, Information Technology, Apache Kafka, Data Pipelines, Docker, Databricks - **Published:** August 19, 2026 - **Apply:** https://job-boards.greenhouse.io/similarweb/jobs/7972320 ## About the Role * Passionate about data. * Holds a BSc degree in Computer Science\Engineering or a related technical field of study. * Has at least 4 years of software or data engineering development experience in one or more of the following programming languages: Python, Java, or Scala. * Has strong programming skills and knowledge of Data Structures, Design Patterns and Object Oriented Programming. * Has good understanding and experience of CI/CD practices and Git. * Excellent communication skills with the ability to provide constant dialog between and within data teams. * Can easily prioritize tasks and work independently and with others. * Conveys a strong sense of ownership over the products of the team. * Is comfortable working in a fast-paced dynamic environment. Advantage: * Has experience with containerization technologies like Docker and Kubernetes. * Experience in designing and productization of complex big data pipelines. * Familiar with a cloud provider (AWS / Azure / GCP). Experience with Big Data technologies and common frameworks such as Spark, Airflow, Kafka, Parquet, Databricks, EMR, Kubernetes. ## Description As a big data engineer developer, you will work at the very core of the company, designing and implementing complex high scale systems to retrieve and analyze data from millions of digital users. Your role as a big data engineer will give you the opportunity to use the most cutting-edge technologies and best practices to solve complex technical problems while demonstrating technical leadership. So, what will you be doing all day? Your role as part of the R&D team means your daily responsibilities may include: * Design and implement complex high scale systems using a large variety of technologies. * You will work in a data research team alongside other data engineers, data scientists and data analysts. Together you will tackle complex data challenges and bring new solutions and algorithms to production. * Contribute and improve the existing infrastructure of code and data pipelines, constantly exploring new technologies and eliminating bottlenecks. * You will experiment with various technologies in the domain of Machine Learning and big data processing. * You will work on a monitoring infrastructure for our data pipelines to ensure smooth and reliable data ingestion and calculation. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market)