> Markdown version of [/jobs/ext/2661125-data-engineer-pxt-central-science](https://www.wearedevelopers.com/jobs/ext/2661125-data-engineer-pxt-central-science). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer, PXT Central Science - **Company:** Amazon.com, Inc. - **Location:** Arlington, VA, United States - **Experience:** Experienced - **Salary:** $152,000.0 - $205,600.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Application Programming Interfaces (APIs), Amazon Web Services, Amazon S3, Data Analysis, Big Data, Code Review, Databases, Data Architecture, Information Engineering, Data Infrastructure, Data Integration, Extract Transform Load (ETL), Data Stores, Data Systems, Graph Database, Apache Hadoop, Apache Hive, Identity and Access Management, Python (Programming Language), Machine Learning, Node.Js, Software Architecture, Scala (Programming Language), Software Engineering, Software Technical Review, Scripting, Apache Spark, Electronic Medical Records, AWS Glue, Non-relational Database, Feature Extraction, Api Design, Software Coding, Software Version Control, Data Pipelines, Amazon Redshift, Programming Languages - **Published:** August 31, 2026 - **Apply:** https://dejobs.org/x/x/E76CF7A3CF424BE78386F4F601651351/job/ ## About the Role * Knowledge of professional software engineering & best practices for full software development life cycle, including coding standards, software architectures, code reviews, source control management, continuous deployments, testing, and operational excellence * 3+ years of data engineering experience * Experience in at least one modern scripting or programming language, such as Python, Java, Scala, or NodeJS * Experience with data modeling, warehousing and building ETL pipelines * Experience with AWS technologies like Redshift, S3, AWS Glue, EMR, Kinesis, FireHose, Lambda, and IAM roles and permissions * Experience with non-relational databases / data stores (object storage, document or key-value stores, graph databases, column-family databases) Preferred Qualifications * Experience with big data technologies such as: Hadoop, Hive, Spark, EMR ## Description PXTCS is looking for a data engineer with expertise in complex data environments. You will be responsible for enhancing our existing data architecture to further standardize metrics and definitions, building and testing new features, developing end-to-end data engineering solutions for complex analytical problems, and collaborating with economists, data scientists, and software engineers to translate data into actionable insights. Specific responsibilities include: * Data Pipeline Development: Design and maintain scalable data pipelines using native AWS services (Glue, EMR, Lambda); build monitoring and error handling for data workflows; optimize performance, reliability, and cost efficiency * Model Productionization & API Development: Develop and maintain APIs and data serving layers that productionize science models for downstream consumption; build batch and real-time inference pipelines * Data Integration & Quality: Build scalable feature extraction and processing frameworks for diverse data types; develop robust data quality and validation checks; create flexible schemas supporting evolving requirements * Cross-team Collaboration: Partner with economics, data science, and software engineering teams to translate analytical requirements into production-ready solutions; participate in technical design reviews and architecture discussions * Analytics & Infrastructure: Maintain layered data systems used by economists and scientists; build automated reporting solutions; work across multiple interconnected AWS accounts with security best practices About the team PXTCS combines economics, behavioral science, statistics, and machine learning to proactively identify mechanisms and process improvements that improve both Amazon's operations and the experience of every Amazonian. Its engineering teams take science-driven insights and models - spanning areas like benefits, compensation, recruiting, voice of employee, management practices, and organizational culture - and turn them into production systems operating at Amazon's scale. PXTCS is an interdisciplinary group where engineering, applied science, and product work side-by-side, and where this team's output directly shapes how Amazon supports its workforce. ## Related Videos - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [How we built an AI-powered code reviewer in 80 hours](https://www.wearedevelopers.com/videos/1511-how-we-built-an-ai-powered-code-reviewer-in-80-hours) - [Stop using Node.js like in 2020! What changed and what you can do today with Node.js](https://www.wearedevelopers.com/videos/100011-stop-using-node-js-like-in-2020-what-changed-and-what-you-can-do-today-with-node-js) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) - [Tips, Techniques, and Common Pitfalls Debugging Kafka](https://www.wearedevelopers.com/videos/838-tips-techniques-and-common-pitfalls-debugging-kafka) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production)