> Markdown version of [/jobs/ext/2120572-data-engineer](https://www.wearedevelopers.com/jobs/ext/2120572-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** SimilarWeb LTD - **Location:** United States (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Airflow, Amazon Web Services, Business Logic, Big Data, Code Review, Data Infrastructure, Raw Data, Data Ingestion, Apache Spark, Software Coding, Amazon Simple Queue Service (SQS), Data Pipelines, Databricks - **Published:** August 19, 2026 - **Apply:** https://job-boards.greenhouse.io/similarweb/jobs/7911963 ## About the Role * Has 5+ years of experience in developing code for big data infrastructure. Proficiency in technologies such as: Databricks, Spark, Airflow, Firehose, SQS, or other similar tools. * Proven experience working with high scale on AWS or any other cloud provider. Experience in architecture and design of large-scale and high performance production systems. * Comfortable taking challenges and learning new technologies. * Excellent communication skills with the ability to provide constant dialog between teams. * Ability to take business requirements and translate them to technical alternatives by performing risk management and evaluating tradeoffs. *All Similarweb offices work in a hybrid model, so you can enjoy the flexibility of working from home with the benefits of building face to face connections with fellow Similarwebbers.* ## Description We're looking for a Senior Data Engineer in the Data Collection Ingest team to contribute to the design and coding of Similarweb's data ingestion services & pipelines. This role involves working as part of a team that handles millions of requests per minute across multiple servers, and is responsible for a wide-range of data pipelines, processing billions of events each day. Why is this role so important at Similarweb? As the most trusted platform for measuring online behavior, millions of people rely on Similarweb's insights daily as the ground truth for their knowledge of the digital world. Producing these insights requires large scale raw data to be ingested reliably in high scale to provide stable signals for analysis. As a Data Engineer you will have the opportunity to perform hands-on work and own Similarweb's raw data ingestion pipeline end-to-end. Your work will have a direct impact on the quality and reliability of our data and the insights that our products are delivering to our customers. So, what will you be doing all day? Your role as Data Engineer of the Ingest Data Collection team means your daily responsibilities may include: * Design, code and manage end-to-end Similarweb's data ingestion pipelines, both online and offline. * Take charge of developing & maintaining modern data infrastructure, while implementing best practices for building data pipelines. * Be responsible for high-scale ingestion services, solving challenges of availability, reliability, and scalability. * Run the production environment by monitoring availability and taking a holistic view of system health and data quality. * Own data infrastructure features from design to production using industry best practices with focus on quality and delivery. * Lead design & decision-making processes of the team. * Solve diverse complex problems of scale, performance and business logic. * Collaborate with product managers and other team leaders to plan, nurture, and implement an efficient and effective development process. * Continuously learn and evaluate new technologies in the everlasting effort to perfect our products * Perform code reviews, evaluate implementations, and provide feedback about potential improvements. * Improve your skills, learn from and mentor top-notch engineers and enrich other team members. * Have lots of fun! ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Data Governance in the Era of AI](https://www.wearedevelopers.com/videos/1622-data-governance-in-the-era-of-ai) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london)