Senior Software Engineer - Production Engineering

One Identity
UK
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
4 years minimum
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Cloud Computing Cloud Engineering Databases Continuous Integration Cursor (Graphical User Interface Elements) Software Debugging DevOps Programming Tools Distributed Systems
+16 more
Log Analysis Node.Js Ruby Runbook Software Engineering Datadog Cloud Platform System GitHub Copilot Grafana Mttr Reliability of Systems Backend Solid Principles BIG-IP Access Policy Manager (APM) Restful APIs GPT

Job description

One Identity’s OneLogin team is seeking Senior Software Engineers who operate with high ownership, strong technical judgment, and a bias for action. These roles are designed for engineers who identify and solve problems proactively - not just execute assigned work. Senior Engineers at OneLogin are expected to operate as T-shaped engineers: deep in a primary domain while remaining capable across the stack (backend, APIs, cloud, and operations). You will play a key role in maintaining product velocity, improving system reliability, and reducing operational burden. This is a hands-on role that includes on-call responsibilities and requires a strong commitment to modern engineering practices, including AI-augmented development. We are looking for engineers who take full ownership - design, build, test, deploy, and operate - rather than handing work off between silos. In this role, you will focus on improving production operability and reliability by debugging complex issues, strengthening observability, and eliminating recurring problems at the source. Responsibilities: Own production operability by debugging complex issues, improving system visibility, and eliminating recurring problems at the source * Own production health for services - from detection through resolution to preventionde

  • Improve mean time to detect (MTTD), mean time to resolve (MTTR), and recurrence rates for issues
  • Identify systemic issues and eliminate recurring problems through code fixes, architecture improvements, and better operational tooling
  • Improve observability across services - logs, metrics, and alerting - for faster diagnosis and resolution
  • Design and improve debugging workflows, runbooks, and internal tooling for engineers
  • Reduce operational burden by making systems easier to understand, operate, and troubleshoot
  • Partner closely with product teams to feed production learnings back into design and development
  • Reduce support and incident load by addressing root causes and improving system design, not just resolving individual issues

Requirements

  • Strong production debugging experience in distributed systems
  • Experience troubleshooting complex, customer-impacting issues under real-world conditions
  • Deep familiarity with observability tools (Datadog or similar - logs, metrics, APM, tracing)
  • Experience improving operational workflows (runbooks, incident response, debugging tooling)
  • Ability to identify patterns across incidents and drive systemic fixes, not just one-off resolutions
  • Experience working across service boundaries (APIs, databases, infrastructure) to diagnose issues
  • Experience using AI tools to analyze logs, incidents, and system behavior at scale to accelerate debugging and root cause identification
  • Strong intuition for system behavior, failure modes, and performance bottlenecks

Qualifications:

Software Engineering

  • 4+ years of software engineering experience with ownership of production systems, reliability, or operational improvements
  • Strong backend development experience (Ruby, Node.js, or similar)
  • Solid understanding of REST APIs, service contracts, and software design principles
  • Experience working across backend services, APIs, and production systems

Cloud & Production Systems

  • Experience building and operating services in AWS or similar cloud environments.
  • Good understanding of distributed systems, cloud-native architecture, and CI/CD.
  • Experience with observability, production debugging, and incident response.

On-Call & Reliability

  • Willingness to participate in a mandatory 24/7 on-call rotation.
  • Experience responding to production incidents and contributing to reliability improvements.

AI-Assisted Development

  • Experience using, or strong interest in, AI-powered development tools (e.g., GitHub Copilot, ChatGPT, Cursor).
  • Ability to evaluate and refine AI-generated code.

About the company

One Identity enables organizations of all sizes to better secure, manage, monitor, protect, and analyze information and infrastructure to help fuel innovation and drive their businesses forward. With team members around the globe, we intend to continue to grow revenues and add value to customers. When you join our team, you will have the opportunity to build and develop products at a scale few others can provide. Our product portfolio serves a large base of customers and we are addressing the strategic imperatives for enterprise businesses. Working with some of the most talented employees the industry has to offer, we provide enhanced career opportunities for team members to learn and grow in a rapidly changing environment. One Identity is an Equal Opportunity Employer and Prohibits Discrimination and Harassment of Any Kind: One Identity is committed to the principle of equal employment opportunity for all employees and to providing employees with a work environment free of discrimination and harassment. All employment decisions at One Identity are based on business needs, job requirements and individual qualifications, without regard to race, color, religion or belief, national, social or ethnic origin, sex (including pregnancy), age, physical, mental or sensory disability, HIV Status, sexual orientation, gender identity and/or expression, marital, civil union or domestic partnership status, past or present military service, family medical history or genetic information, family or parental status, or any other status protected by the laws or regulations in the locations where we operate. One Identity will not tolerate discrimination or harassment based on any of these characteristics. One Identity encourages applicants of all ages.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on uk.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:14 min

Exploring internal AI product initiatives and global engineering roles

Maria Apazoglou · Coffee With Developers

50 sec

Why developer happiness matters in web frameworks

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

3:08 min

Aligning engineering processes with core business impact metrics

Chris Riley · WWC 2021

40 sec

Generative pre-trained transformer models powering code completions

lgonta lgonta +1 · WWC 2024

3:30 min

Falling in love with Ruby and creating Basecamp

David Heinemeier Hansson David Heinemeier Hansson +1 · Coffee With Developers

3:07 min

Establishing service level agreements directly for internal platforms

Pawel Piwosz · LIVE

Videos

See all

Related articles

See all