Staff AI Data Engineer

PlayStation
Seattle, WA, United States
13 days ago
Apply on job-boards.greenhouse.io
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
7 years minimum
Compensation
$177,300.0 - $265,900.0
Working hours
Regular working hours

Tech stack

Query Performance Application Programming Interfaces (APIs) Artificial Intelligence Airflow Amazon Web Services Data Analysis Architectural Patterns Automation of Tests BigQuery Cloud Computing Cloud Engineering Code Review
+34 more
Continuous Integration Data as a Services Information Engineering Data Governance Data Infrastructure Extract Transform Load (ETL) Data Warehousing Relational Databases DevOps Distributed Data Store Information Lifecycle Management Python (Programming Language) PostgreSQL Machine Learning Automation of Marketing Operational Databases Search Technologies Software Deployment Unstructured Data Cloud Platform System Retrieval-Augmented Generation Large Language Models Snowflake Apache Spark Backend Git Build Management Containerization AI Platforms Apache Kafka Data Management Terraform Data Pipelines Amazon Redshift

Job description

We’re looking for a Staff AI Data Engineer to design and build the data infrastructure, pipelines, and services that power AI and machine learning workflows across PlayStation Studios. You’ll own systems that move, transform, organize, and serve data for LLM applications, agentic workflows, analytics, experimentation, and production AI services.

This is a hands-on Staff-level engineering role spanning data engineering, backend systems, cloud infrastructure, and applied AI. You’ll work across the stack to build reliable data platforms and services, establish scalable architectural patterns, and help teams turn studio data into useful, production-ready AI capabilities.

What You’ll Be Doing

  • Design, build, and own scalable data infrastructure supporting LLM, agentic, machine learning, analytics, and experimentation workflows.

  • Build and operate reliable batch and event-driven data pipelines for ingestion, transformation, enrichment, and delivery across varied studio data sources.

  • Develop ETL/ELT pipelines that feed data warehouses and curated datasets for analysts, analytics engineers, ML engineers, and downstream applications.

  • Design data models, APIs, and storage patterns that make structured and unstructured data easy to discover, access, and use across AI-powered systems.

  • Build data-backed services and platform capabilities for LLM applications, including retrieval-augmented generation, embeddings, vector search, tool integrations, and context retrieval.

  • Develop backend services and reusable components in Python that support production AI and data workflows.

  • Design and operate cloud infrastructure using AWS and/or GCP, Infrastructure as Code, containerized workloads, and modern CI/CD practices.

  • Establish standards for data quality, lineage, observability, testing, schema evolution, reliability, security, and operational ownership.

  • Evaluate new technologies and architectural approaches across data engineering and the rapidly evolving LLM and agent ecosystem, and determine where they provide practical value.

  • Partner with AI/ML engineers, software engineers, analysts, analytics engineers, researchers, and studio teams to translate ambiguous requirements into scalable technical solutions.

  • Provide technical leadership across projects, influence architecture and engineering standards, mentor other engineers, and help shape the long-term direction of the AI Engineering data platform.

  • Produce clear technical documentation, architectural guidance, examples, and reusable patterns that allow solutions to scale across teams and studios., * Disabled Veteran: a veteran of the U.S. military who is entitled to compensation (or who would be entitled to compensation if not for the receipt of military retired pay) under the administration of the Secretary of Veterans Affairs and/or a person who was discharged or released from active duty because of a service-connected disability.
  • Recently Separated Veteran: a veteran who has discharged or released from active duty in the U.S. military within the last three years.
  • Armed Forces Service Medal Veteran: a veteran who, while serving on active duty in the U.S. military, participated in a U.S. military operation for which an Armed Forces Service Medal was awarded pursuant to Executive Order 12985.

Requirements

  • You have at least seven years of experience building and operating production data platforms, backend systems, or distributed data-intensive applications.

  • You are highly proficient in Python and have experience building maintainable production software, not just scripts or notebooks.

  • You have designed and operated production data pipelines, including ingestion, transformation, orchestration, monitoring, failure recovery, and data-quality validation.

  • You have strong experience with relational databases and analytical data platforms such as PostgreSQL, Redshift, Snowflake, BigQuery, or similar systems.

  • You are experienced with cloud-native architecture on AWS, GCP, or another major cloud platform and understand networking, security, storage, compute, and managed data services.

  • You have hands-on experience with Infrastructure as Code such as Terraform and with modern CI/CD and DevOps practices.

  • You understand data modeling, schema design, query performance, partitioning, data lifecycle management, and the tradeoffs between transactional, analytical, and specialized storage systems.

  • You have experience designing APIs, services, or other programmatic interfaces for accessing and operating on data.

  • You are comfortable working in collaborative Git-based development environments with code review, automated testing, and production deployment workflows.

  • You communicate clearly, write strong technical documentation, and can translate broad or ambiguous problems into pragmatic, maintainable systems.

  • You operate effectively at Staff level: independently driving architecture and execution, influencing technical direction across teams, and raising engineering standards beyond your immediate projects.

Nice to Have

  • You have built production systems using LLMs, agentic workflows, retrieval-augmented generation, embeddings, or vector databases.

  • You have experience with LLM application infrastructure and tooling such as MCP, tool calling, evaluation systems, prompt/context management, or model observability.

  • You have worked with orchestration and distributed data-processing technologies such as Airflow, Prefect, Dagster, Spark, Kafka, or similar systems.

  • You have experience working with both structured and unstructured data, including documents, source code, telemetry, logs, media metadata, or other large-scale content.

  • You have built internal data platforms, self-service developer platforms, or reusable infrastructure used by multiple engineering teams.

  • You are familiar with data governance, privacy, security, access controls, lineage, retention, and responsible-AI considerations for enterprise data.

  • You have experience supporting ML training, evaluation, feature generation, or other machine learning data workflows.

At SIE, we consider several factors when setting each role’s base pay range, including the competitive benchmarking data for the market and geographic location.

Please note that the base pay range may vary in line with our hybrid working policy and individual base pay will be determined based on job-related factors which may include knowledge, skills, experience, and location., Under U.S. law, you are considered to have a disability if you have a physical or mental impairment or medical condition that substantially limits a major life activity, or if you have a history or record of such an impairment or medical condition.

Identifying yourself as an individual with a disability is voluntary, and we hope that you will choose to do so. Your answer will be maintained confidentially and will not be seen by selecting officials or anyone else involved in making personnel decisions, nor will it be shared with our accommodations team . Completing the form will not negatively impact you in any way, regardless of whether you have self-identified in the past.

Benefits & conditions

Select… Will you now, or in the future, require sponsorship to work in the United States?* Select… Will you need relocation assistance to work at this role’s specified location?* Select… Are you related to, or in a close personal relationship with, anyone who currently works for SIE or any SIE-affiliated studios? (This includes spouses, domestic partners, and significant others.)* Select… If yes, please state their name, the department or studio they work for, and their job title (if you know it). By selecting “Yes”, I am certifying that, to the best of my knowledge, the information I have provided in this employment application is true and correct.* Select… LinkedIn Profile URL Portfolio Website

About the company

Sony Interactive Entertainment isn’t just the Best Place to Play - it’s also the Best Place to Work. Sony Interactive Entertainment (SIE) is the company behind the PlayStation brand. As a subsidiary of Sony Group Corporation, we’re part of a proud legacy of innovation and excellence. SIE is a dynamic technology company, delivering cutting-edge hardware and network services to more than 100 million people and an entertainment leader, home to some of the most beloved and recognizable intellectual properties (IP) in the world. Our role at SIE is to create and nurture the experiences under the PlayStation brand, a name synonymous with entertainment excellence and creativity.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on job-boards.greenhouse.io
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

2:37 min

Optimizing technical profiles for AI sourcing and recruitment

Mina Golesorkhi Mina Golesorkhi · World Congress 2026 Europe

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

Videos

See all

Related articles

See all