Software Engineer II Data Reliability & Automation (APIs)

PlayStation
San Mateo, CA, United States
7 days ago
Apply on www.indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Compensation
$150,100.0 - $225,100.0
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Automation of Tests Cloud Computing Code Review Databases Data as a Services Data Integrity Linux Amazon DynamoDB Failover
+29 more
Fault Tolerance Design of User Interfaces NoSQL Redis Reliability Engineering E2e Testing Software Engineering SQL Databases Data Streaming Systems Integration Management of Software Versions Openapi Data Logging Aerospike Google Cloud Amazon ElastiCache Caching Reliability of Systems Database Performance Backend Containerization Kubernetes Infrastructure Automation Frameworks Information Technology Cassandra Performance Monitor Apache Kafka Api Design Terraform

Job description

This role combines backend software development with infrastructure automation and operational engineering. You’ll develop Go services, APIs, Terraform modules, and automation for provisioning and managing databases, streaming platforms, and caching services such as Cassandra, Aurora, Aerospike, Kafka, and Redis.

You’ll also contribute to enabling self-service experiences that translate developer intent into validated infrastructure configurations and provisioning workflows. Working alongside senior engineers and platform teams, you’ll help ensure these capabilities are reliable, secure, observable, and safe to operate at scale., * Develop and maintain software, APIs, and automation for provisioning and managing database, streaming, and caching infrastructure.

  • Build and maintain reusable Terraform modules and Infrastructure as Code workflows for cloud, Kubernetes, and data-platform services.
  • Develop backend services and platform tooling in Go to support infrastructure provisioning, configuration, and lifecycle management.
  • Automate operational activities such as deployment, configuration, scaling, upgrades, backup, recovery, failover, and decommissioning.
  • Help operate and improve SQL, NoSQL, streaming, and caching platforms at scale, including technologies such as Cassandra, Aurora, Aerospike, Kafka/MSK, Redis, DynamoDB, and ElastiCache.
  • Contribute to the availability, scalability, performance, and reliability of stateful data services
  • Help build self-service and agent-assisted experiences that translate developer intent into structured infrastructure specifications and Terraform-based execution plans.
  • Build and enhance observability using metrics, logs, traces, dashboards, and actionable alerts.
  • Participate in on-call rotations and incident response for database, caching, streaming, and supporting platform services.
  • Troubleshoot infrastructure and application issues involving database performance, capacity, connectivity, replication, storage, and resource utilization.
  • Contribute to root-cause analysis and implement automated or systemic fixes that prevent recurring incidents.
  • Help define and measure SLIs, SLOs, error budgets, and operational health indicators for platform services.
  • Create automated tests for APIs, Terraform modules, infrastructure plans, operational tooling, and failure scenarios.
  • Collaborate with engineering, platform, security, and operations teams to deliver reliable and scalable data services.
  • Maintain technical documentation, API specifications, operational playbooks, runbooks, and onboarding materials.
  • Participate in code reviews, design discussions, capacity planning, and continuous improvement initiatives.

Requirements

  • Bachelor’s or Master’s degree in Computer Science or a related field, or equivalent practical experience.
  • 3+ years of experience in software engineering, database reliability engineering, site reliability engineering, platform engineering, infrastructure automation, or a related field.
  • Experience developing production software, backend services, or infrastructure tooling in Go.
  • Hands on experience using Terraform to provision and manage cloud infrastructure.
  • Experience developing APIs, including interface design, validation, error handling, testing, and versioning.
  • Working knowledge of one or more SQL, NoSQL, streaming, or caching technologies, such as Cassandra, Aurora, Aerospike, Kafka, AWS MSK, Redis, DynamoDB, or ElastiCache.
  • Understanding of database and distributed-system concepts such as replication, partitioning, consistency, availability, fault tolerance, backup, and recovery.
  • Familiarity with the operational requirements of stateful systems, including capacity planning, performance monitoring, scaling, upgrades, backup and recovery, and incident troubleshooting.
  • Experience with AWS or GCP, including familiarity with cloud networking, identity, compute, storage, and managed data services.
  • Working knowledge of Kubernetes and experience deploying or operating containerized workloads.
  • Understanding of Linux, networking, storage systems, and common troubleshooting techniques.
  • Familiarity with observability practices, including metrics, structured logging, tracing, alerting, and dashboards.
  • Experience writing unit, integration, and automated end-to-end tests.
  • Strong problem-solving and analytical skills, with an interest in infrastructure automation and distributed-system reliability.
  • Strong written and verbal communication skills and the ability to collaborate effectively across teams., * Experience building internal developer platforms, self-service infrastructure, or infrastructure APIs.
  • Experience using AI/ML or generative AI to improve infrastructure automation, operational workflows, incident management, observability, or developer productivity.
  • Familiarity with agent-assisted workflows, structured outputs, tool-calling integrations, approval controls, or policy-based infrastructure automation.
  • Relevant AWS, GCP, Kubernetes, Terraform, database, or streaming-platform certifications.

About the company

Sony Interactive Entertainment is a Fair Chance employer and qualified applicants with arrest and conviction records will be considered for employment.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

3:55 min

Demonstrating semantic routing thresholds with the Redis vector library

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

5:00 min

Exploring the specific workplace responsibilities of staff software engineers

Jan Giacomelli · LIVE

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

Videos

See all

Related articles

See all