> Markdown version of [/jobs/ext/1250962-data-engineer-hybrid](https://www.wearedevelopers.com/jobs/ext/1250962-data-engineer-hybrid). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - Hybrid - **Company:** SmartIMS Inc. - **Location:** Arlington, VA, United States - **Salary:** $83,200.0 - $166,400.0 - **Contract:** Permanent contract - **Skills:** Clean Code Principles, Big Data, Code Review, Extract Transform Load (ETL), Data Systems, Database Design, Database Testing, Distributed Computing Environment, Apache Hadoop, Hadoop Distributed File System, Apache Hive, SQL Databases, Enterprise Data Management, Software Organization, Data Processing, Apache Yarn, Sql Optimization, Apache Spark, Pyspark, Information Technology, Spark Streaming, Data Management, Software Version Control, Data Pipelines - **Published:** July 12, 2026 - **Apply:** https://www.careerjet.com/jobad/us10426e8680cae973936d376a497a4b86 ## About the Role * Experience as a Data Engineer or in a similar data-focused engineering role * Strong expertise in writing and optimizing SQL queries for large datasets * Hands-on experience with Apache Spark, including PySpark, Spark SQL, and Spark Streaming * Experience working with Hadoop ecosystem technologies such as HDFS, Hive, and YARN * Strong understanding of ETL frameworks and data pipeline development * Knowledge of distributed data processing and big data architectures * Understanding of data modeling concepts and database design principles * Experience working with Python for data engineering and automation tasks * Ability to analyze, troubleshoot, and resolve complex data issues independently * Knowledge of data testing, validation, and quality assurance practices * Strong verbal and written communication skills with the ability to collaborate with technical and non-technical stakeholders * Experience with version control, code reviews, and software development best practices * Ability to work effectively in a collaborative, fast-paced environment * Bachelor's degree in Engineering, Mathematics, Finance, Business, Computer Science, or a related quantitative field, or equivalent practical experience ## Description As a Data Engineer, you will support the design, development, and maintenance of enterprise data platforms and large-scale data processing solutions. You will be responsible for building and optimizing data pipelines, developing ETL processes, and ensuring the availability, reliability, and quality of data across business-critical systems. This role requires expertise in big data technologies, SQL optimization, and distributed data processing, along with the ability to collaborate with cross-functional teams to deliver scalable and efficient data solutions., * Design, implement, and maintain enterprise ETL processes and data pipelines * Develop scalable and efficient code to process, transform, and deliver large datasets * Build and optimize data pipelines using Apache Spark, Hadoop, and related big data technologies * Collaborate with engineering and analytics teams to solve complex data challenges and maintain data quality * Support the delivery of accurate and actionable data solutions for business stakeholders * Design and manage distributed data processing workflows and orchestration processes * Develop and optimize SQL queries for large-scale data retrieval, transformation, and analysis * Participate in data modeling and database design initiatives to support scalable solutions * Monitor, troubleshoot, and resolve data processing and pipeline issues * Automate routine data management tasks and improve operational efficiency * Apply testing and validation practices to ensure data accuracy, consistency, and reliability * Participate in code reviews and follow development best practices and version control standards * Build strong working relationships with internal teams and business stakeholders * Ensure compliance with organizational policies, standards, and regulatory requirements ## Related Videos - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Are Code Reviews Worth It? Insights from 16 Years of Review Data](https://www.wearedevelopers.com/videos/1135-are-code-reviews-worth-it-insights-from-16-years-of-review-data) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Top-Paying Tech Jobs (with Salaries)](https://www.wearedevelopers.com/magazine/372-top-paying-tech-jobs-with-salaries) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer)