Data Engineer, PXT Central Science

Amazon.com, Inc.
Arlington, VA, United States
6 days ago
Apply on dejobs.org
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Compensation
$152,000.0 - $205,600.0
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Application Programming Interfaces (APIs) Amazon Web Services Amazon S3 Data Analysis Big Data Code Review Databases Data Architecture Information Engineering Data Infrastructure Data Integration
+26 more
Extract Transform Load (ETL) Data Stores Data Systems Graph Database Apache Hadoop Apache Hive Identity and Access Management Python (Programming Language) Machine Learning Node.Js Software Architecture Scala (Programming Language) Software Engineering Software Technical Review Scripting Apache Spark Electronic Medical Records AWS Glue Non-relational Database Feature Extraction Api Design Software Coding Software Version Control Data Pipelines Amazon Redshift Programming Languages

Job description

PXTCS is looking for a data engineer with expertise in complex data environments. You will be responsible for enhancing our existing data architecture to further standardize metrics and definitions, building and testing new features, developing end-to-end data engineering solutions for complex analytical problems, and collaborating with economists, data scientists, and software engineers to translate data into actionable insights. Specific responsibilities include:

  • Data Pipeline Development: Design and maintain scalable data pipelines using native AWS services (Glue, EMR, Lambda); build monitoring and error handling for data workflows; optimize performance, reliability, and cost efficiency

  • Model Productionization & API Development: Develop and maintain APIs and data serving layers that productionize science models for downstream consumption; build batch and real-time inference pipelines

  • Data Integration & Quality: Build scalable feature extraction and processing frameworks for diverse data types; develop robust data quality and validation checks; create flexible schemas supporting evolving requirements

  • Cross-team Collaboration: Partner with economics, data science, and software engineering teams to translate analytical requirements into production-ready solutions; participate in technical design reviews and architecture discussions

  • Analytics & Infrastructure: Maintain layered data systems used by economists and scientists; build automated reporting solutions; work across multiple interconnected AWS accounts with security best practices

About the team

PXTCS combines economics, behavioral science, statistics, and machine learning to proactively identify mechanisms and process improvements that improve both Amazon’s operations and the experience of every Amazonian. Its engineering teams take science-driven insights and models - spanning areas like benefits, compensation, recruiting, voice of employee, management practices, and organizational culture - and turn them into production systems operating at Amazon’s scale. PXTCS is an interdisciplinary group where engineering, applied science, and product work side-by-side, and where this team’s output directly shapes how Amazon supports its workforce.

Requirements

  • Knowledge of professional software engineering & best practices for full software development life cycle, including coding standards, software architectures, code reviews, source control management, continuous deployments, testing, and operational excellence
  • 3+ years of data engineering experience
  • Experience in at least one modern scripting or programming language, such as Python, Java, Scala, or NodeJS
  • Experience with data modeling, warehousing and building ETL pipelines
  • Experience with AWS technologies like Redshift, S3, AWS Glue, EMR, Kinesis, FireHose, Lambda, and IAM roles and permissions
  • Experience with non-relational databases / data stores (object storage, document or key-value stores, graph databases, column-family databases)

Preferred Qualifications

  • Experience with big data technologies such as: Hadoop, Hive, Spark, EMR

Benefits & conditions

The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits .

USA, CA, San Francisco - 152,000.00 - 205,600.00 USD annually

USA, VA, Arlington - 132,100.00 - 178,800.00 USD annually

USA, WA, Bellevue - 132,100.00 - 178,800.00 USD annually

USA, WA, Seattle - 132,100.00 - 178,800.00 USD annually

About the company

Amazon’s People Experience and Technology Central Science (PXTCS) team uses economics, behavioral science, statistics, machine learning, and Generative AI to proactively identify mechanisms and process improvements that simultaneously improve Amazon and the lives, well-being, and value of work for Amazonians. PXTCS is an interdisciplinary team that combines the talents of science, engineering, and UX to build and deliver solutions that measurably achieve this goal - at a scale that touches over 1.5 million Amazonians worldwide.

As a Data Engineer on PXTCS, you’ll work side by side with economists, data scientists, software engineers, and applied scientists turning leading-edge ML and Generative AI models into reliable, scalable production systems.

This is a rare chance to see your code directly shape how Amazon supports its workforce, spanning areas like benefits, compensation, recruiting, voice of employee, management practices, and organizational culture. We offer opportunities for builders to build and make history!

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dejobs.org
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

55 sec

Validating data processing architectures via containerized events

Modood Alvi · World Congress 2025

45 sec

Working securely with Node.js path application programming interfaces

Sonya Moisset · World Congress 2023

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

Videos

See all

Related articles

See all