> Markdown version of [/jobs/ext/2148594-emr-spark-sql-job-query-tuning-engineer](https://www.wearedevelopers.com/jobs/ext/2148594-emr-spark-sql-job-query-tuning-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # EMR/Spark SQL & Job Query-Tuning Engineer - **Company:** Amazon.com, Inc. - **Location:** Seattle, WA, United States - **Salary:** $91,000.0 - $152,000.0 - **Contract:** Permanent contract - **Skills:** Big Data, Catalyst (Software), Cloud Computing, Directed Acyclic Graph (Directed Graphs), Information Engineering, Serialization, Distributed Computing Environment, Distributed Systems, Apache Hive, Performance Tuning, Query Optimization, SQL Databases, Parquet, Data Processing, Apache Spark, Pyspark, Information Technology, Integration Frameworks, AWS Data Analytics, Amazon Elastic Mapreduce (EMR) - **Published:** August 20, 2026 - **Apply:** https://www.careerjet.com/jobad/us73ef2ad6a48cd6079bda62cba22b7809 ## About the Role * Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field, or equivalent professional experience. * Strong hands-on experience with PySpark, Spark SQL, Hive SQL, and Amazon EMR. * Proven expertise in Spark performance tuning, query optimization, and execution plan analysis. * Deep understanding of Spark internals, distributed computing concepts, and data processing frameworks. * Experience optimizing large-scale data workloads and addressing memory, partitioning, and shuffle-related challenges. * Proficiency in analyzing Spark UI metrics and troubleshooting performance issues. * Strong problem-solving skills and ability to work in a fast-paced environment. Preferred Qualifications * Knowledge of Adaptive Query Execution (AQE), Catalyst Optimizer, and Tungsten Engine. * Experience working with Parquet, ORC, and other columnar data formats. * Familiarity with AWS data services and cloud-based big data platforms. * Experience supporting enterprise-scale analytics or data engineering environments. ## Description STAFFXPERT LLC is seeking an EMR/Spark SQL & Job Query-Tuning Engineer on behalf of our client in Seattle, WA. This role is ideal for a Spark performance expert who specializes in optimizing PySpark, Spark SQL, and Hive workloads within large-scale data environments. The successful candidate will focus on improving application-level performance by tuning queries, optimizing execution plans, reducing memory consumption, minimizing shuffle operations, and enhancing overall job efficiency on Amazon EMR. Key Responsibilities * Optimize and tune PySpark, Spark SQL, and Hive SQL jobs to improve performance, scalability, and resource utilization. * Analyze Spark execution plans and DAGs to identify and resolve performance bottlenecks. * Design and implement efficient partitioning, caching, and data processing strategies. * Reduce shuffle overhead, spill events, executor memory pressure, and job execution times. * Optimize join strategies, including broadcast joins, sort-merge joins, and other Spark execution techniques. * Troubleshoot data skew, serialization issues, and distributed processing inefficiencies. * Collaborate with data engineering and platform teams to improve workload performance and reliability. * Monitor, benchmark, and continuously enhance large-scale data processing jobs. * Recommend and implement best practices for Spark and EMR performance optimization., Hi, Role: EMR/Spark SQL & Job Query-Tuning Engineer Location:Redmond,WA skills Required: Application/code-level optimization of PySpark/Spark/Hive SQL jobs; query tuning, D… + 20 hours ago + Apply easily + ## Related Videos - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Parquet, Delta, Iceberg & Ducklake - An introduction for developers](https://www.wearedevelopers.com/videos/100075-parquet-delta-iceberg-ducklake-an-introduction-for-developers) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [OLAP for AI Applications and why you should care](https://www.wearedevelopers.com/videos/100212-olap-for-ai-applications-and-why-you-should-care) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) ## Related Articles - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production)