> Markdown version of [/jobs/ext/2113092-data-engineer](https://www.wearedevelopers.com/jobs/ext/2113092-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # data engineer - **Company:** SimilarWeb LTD - **Location:** United States (Remote available) - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Airflow, Amazon Web Services, Big Data, C Sharp (Programming Language), Databases, Fault Tolerance, Python (Programming Language), PostgreSQL, Raw Data, Redis, Software Engineering, Large Language Models, Apache Spark, Pyspark, Information Technology, Machine Learning Operations, Databricks, Web Api, Golang, Programming Languages, Microservices - **Published:** August 19, 2026 - **Apply:** https://job-boards.greenhouse.io/similarweb/jobs/7556162 ## About the Role * Holds a BSc degree in Computer Science or equivalent practical experience. * You love building robust, fault-tolerant, and scalable systems and products * You are a go-getter and a team player with a sense of ownership. * Has at least 3+ years of server-side software development experience in one or more general-purpose programming languages (C#, Go, Python, etc.) * Experience building large-scale web APIs: advantage for working with Microservices architecture, AWS, and databases (Redis, PostgreSQL, Firebolt) * Familiarity with Big Data technologies: A familiarity with Spark, Databricks, and Airflow is a big advantage. * Worked in a cloud environment such as AWS or GCP, and is familiar with its different services. * Familiarity with ML pipelines and applications * Familiarity with LLM tools and frameworks *All Similarweb offices work in a hybrid model, so you can enjoy the flexibility of working from home with the benefits of building face to face connections with fellow Similarwebbers.* ## Description We're looking for a Data Engineer to join our Data Labs (DL) department, which specializes in professional services for our super-premium customers. This role will report to our DL (Datalabs) Data Science Team Manager in the R&D. Why is this role so important at Similarweb? Similarweb is a data-focused company, and our unique AI and machine learning capabilities are the center of our business. As part of this role, you will create and support a complex data model pipeline that helps analyze the petabytes of data we receive from various sources, and research and develop new features and capabilities for our product solutions As a data engineer in the Datalabs team, you will work on the very core of the company. Part of your role will be to create processes that help turn raw data into usable metrics and leverage AI models and statistical algorithms to support out-of-the-box requests from customers who want custom data labs. The Datalabs department's business-oriented nature also means you will be supporting a team of analysts and data scientists who interact directly with customers. Together with them, you will translate the voice of these customers into best-in-class data labs. So, what will you be doing all day? * Building and maintaining our big-data pipelines * Take a major part in designing and implementing complex high-scale systems using a large variety of technologies * Be part of a team with smart and motivated engineers, and data scientists, to collaborate on the planning, development, and maintenance of our products * Implement solutions in the AWS cloud environment, and work in Databricks with PySpark ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Reducing LLM Calls with Vector Search Patterns - Raphael De Lio (Redis)](https://www.wearedevelopers.com/videos/1714-reducing-llm-calls-with-vector-search-patterns-raphael-de-lio-redis) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) - [Scoring 2000 Products per Request: Performance Pitfalls in Golang](https://www.wearedevelopers.com/videos/2073-scoring-2000-products-per-request-performance-pitfalls-in-golang) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market) - [Top-Paying Tech Jobs (with Salaries)](https://www.wearedevelopers.com/magazine/372-top-paying-tech-jobs-with-salaries)