> Markdown version of [/jobs/ext/2553766-python-pyspark-developer](https://www.wearedevelopers.com/jobs/ext/2553766-python-pyspark-developer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Python PySpark Developer - **Company:** Hexaware Technologies - **Location:** McLean, VA, United States - **Contract:** Permanent contract - **Skills:** Clean Code Principles, Java (Programming Language), Application Programming Interfaces (APIs), Application Integration Architecture, Unit Testing, Big Data, Software Quality, Code Review, Databases, Data Validation, Information Engineering, Data Infrastructure, Data Integration, Extract Transform Load (ETL), Data Transformation, Data Structures, Relational Databases, Distributed Computing Environment, Human-Computer Interaction, Apache Struts, Python (Programming Language), Performance Tuning, Standard Sql, Software Engineering, Web Application Frameworks, Data Processing, Enterprise Software Applications, Git, AngularJS, Pyspark, Integration Tests, Information Technology, Software Version Control, Data Pipelines - **Published:** August 31, 2026 - **Apply:** https://www.dice.com/job-detail/1a624d0f-8f02-4118-b1f2-395951de4604 ## About the Role Bachelor s degree in Computer Science, Engineering, Information Technology, or a related discipline. - Strong hands-on experience with Python development. - Practical experience with PySpark and distributed data processing. - Experience developing ETL or data engineering pipelines. - Good understanding of data structures, transformation logic, and performance optimization. - Experience working with SQL and relational databases. - Experience with version control systems, preferably Git. - Ability to write clean, modular, and maintainable code. - Understanding of software development life cycle practices. - Experience with unit testing, integration testing, and code quality practices. - Strong problem-solving and analytical skills. - Ability to work effectively in a collaborative enterprise development environment. - Good communication skills and the ability to understand business requirements. ## Description The primary focus of this role will be on Python development, PySpark-based data processing, ETL development, and data integration. The successful candidate will also work with application and user-interface components. Experience with Angular or legacy Java web frameworks such as Struts is beneficial, but candidates with strong Python and PySpark capabilities who are willing to learn these technologies will also be considered., Develop, enhance, and maintain data processing applications using Python and PySpark. - Design and implement scalable ETL and data transformation workflows. - Process and analyze large datasets using distributed computing frameworks. - Integrate data from databases, files, APIs, and enterprise applications. - Write reusable, maintainable, and well-tested Python code. - Troubleshoot data pipelines, application issues, and production defects. - Perform data validation, reconciliation, and quality checks. - Collaborate with data engineers, application developers, business analysts, and QA teams. - Support the development of reporting, monitoring, and operational interfaces. - Contribute to application integration between modern data platforms and existing enterprise systems. - Learn and support Angular-based front-end components as required. - Assist with the maintenance and gradual enhancement of legacy Struts-based applications. - Participate in code reviews, technical documentation, testing, and deployment activities. ## Related Videos - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) ## Related Articles - [The 13 Best Python Libraries for Developers in 2025](https://www.wearedevelopers.com/magazine/371-the-13-best-python-libraries-for-developers-in-2025) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Everything a Developer Needs to Know About MCP with Neo4j](https://www.wearedevelopers.com/magazine/604-everything-a-developer-needs-to-know-about-mcp-with-neo4j) - [7 good reasons why you should learn Python in 2021](https://www.wearedevelopers.com/magazine/19-7-good-reasons-why-you-should-learn-python-in-2021)