> Markdown version of [/jobs/ext/660205-site-reliability-engineer](https://www.wearedevelopers.com/jobs/ext/660205-site-reliability-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Site Reliability Engineer - **Company:** TechSpace Solutions Inc. - **Location:** Cincinnati, OH, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Application Programming Interfaces (APIs), Amazon Web Services, Cloud Computing, Databases, Data Systems, Software Debugging, Distributed Systems, Payment Systems, Performance Tuning, Ruby on Rails, Reliability Engineering, Runbook, Software Engineering, SQL Databases, Datadog, Grafana, Information Technology, Splunk, New Relic (SaaS), Service Stack, Microservices - **Published:** June 26, 2026 - **Apply:** https://www.dice.com/job-detail/314e66ea-00b3-4743-84ac-aa85f71b1187 ## About the Role * 10+ years of experience in Software Engineering, Production Engineering, SRE, or Distributed Systems. * Strong experience debugging production issues end-to-end (application, infrastructure, data, and dependencies). Hands-on experience with: * AWS and cloud-native environments * Ruby on Rails and/or Java * APIs, Microservices, and Distributed Systems * SQL and database troubleshooting * Observability tools such as Splunk, Datadog, New Relic, etc. Deep understanding of: * System behavior in production * Fault isolation and troubleshooting * Performance optimization and resiliency patterns * Excellent communication and stakeholder management skills. * Ability to work effectively during incidents and high-pressure situations. Preferred Qualifications: * Experience in Payments, FinTech, Banking, or other regulated environments. * Experience building and operating large-scale, high-availability platforms. * Bachelor's degree in Computer Science, Engineering, or equivalent practical experience. ## Description * Client is looking for an enterprise-grade embedded finance platform enabling organizations to build, launch, and scale compliant banking, payments, and lending solutions. * We are seeking a Principal Software Engineer to join our Production Engineering team. This is a hands-on technical leadership role focused on operating, debugging, and improving highly distributed, mission-critical payment systems. The ideal candidate thrives in complex production environments and enjoys solving deep technical challenges across applications, infrastructure, and data systems., * Lead production triage and incident response across APIs, payment systems, distributed services, infrastructure, and databases. * Diagnose and resolve complex production issues spanning code, infrastructure, data, and third-party dependencies. * Partner with engineering teams to implement permanent fixes and improve platform reliability. * Design and implement monitoring, alerting, automation, and operational tooling. * Improve system observability, resiliency, and debuggability. * Work across a mixed technology stack including Ruby on Rails, Java, AWS, APIs, and SQL databases. * Develop runbooks and diagnostic workflows for operational excellence. * Mentor engineers and influence best practices across engineering and SRE teams. * Participate in architectural discussions to build highly reliable and scalable systems. ## Related Videos - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [Debugging in the Dark](https://www.wearedevelopers.com/videos/1658-debugging-in-the-dark) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Software Engineering Social Connection: Yubo’s lean approach to scaling an 80M-user infrastructure](https://www.wearedevelopers.com/videos/1583-software-engineering-social-connection-yubo-s-lean-approach-to-scaling-an-80m-user-infrastructure) - [Leading with Reliability: Applying SRE Principles to Build Stronger Engineering Organizations](https://www.wearedevelopers.com/videos/100185-leading-with-reliability-applying-sre-principles-to-build-stronger-engineering-organizations) ## Related Articles - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Best Countries for Software Engineers](https://www.wearedevelopers.com/magazine/267-best-countries-for-software-engineers) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs)