> Markdown version of [/jobs/ext/1316144-software-engineer-sre](https://www.wearedevelopers.com/jobs/ext/1316144-software-engineer-sre). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Software Engineer - SRE - **Company:** OneTrust - **Location:** United States - **Experience:** Expert - **Salary:** $116,475.0 - $174,713.0 - **Contract:** Permanent contract - **Skills:** Query Performance, Java (Programming Language), Artificial Intelligence, Amazon Web Services, Software Applications, Computing Platforms, Microsoft Azure, Databases, Distributed Systems, Java Virtual Machine (JVM), Python (Programming Language), Machine Learning, NoSQL, Pattern Recognition, Reliability Engineering, Prometheus, Ruby, Software Engineering, SQL Databases, Datadog, Data Logging, Google Cloud, Cloud Platform System, Large Language Models, Grafana, Prompt Engineering, Mttr, Reliability of Systems, Database Performance, Kubernetes, Information Technology, Terraform, Pagerduty, Jenkins, Microservices - **Published:** July 17, 2026 - **Apply:** https://www.dice.com/job-detail/3ba2e696-d1a5-42fa-b02d-97a1c61b8143 ## About the Role * Bachelor's degree in computer science, Engineering, or related technical or business field * 4+ yrs. of application development experience with Java or other equivalent language * Experience with Spring environment * Experience in cloud-based infrastructure (Azure, AWS, Google Cloud Platform, etc.) * Experience with the factors that affect software application performance at different levels. These factors include database performance, network performance, CPU utilization, JVM tuning, memory analysis, thread management, and query performance. * A knowledge of the importance of centralizing logging, metrics dashboards, and alerting. Able to articulate about some of the tools used for these tasks * A good awareness of databases (ideally SQL/NoSQL) * Hands-on experience with observability tools (Datadog, Prometheus, Grafana, etc.) * Knowledge with CI/CD pipelines and infrastructure-as-code (Terraform, Helm, jenkins, gitlab) * Build and operate AI-assisted incident response systems (root cause analysis, log summarization, anomaly triage) * Develop or integrate LLM-based tools to reduce MTTR and improve alert quality * Apply machine learning techniques for anomaly detection, capacity prediction, or failure pattern analysis * Experience deploying AI systems in production (not just experimentation) * Knowledge with vector databases, embeddings, or RAG architectures for operational intelligence * Well-developed insight of prompt engineering and evaluation of LLM outputs in the reliability workflow * Kubernetes and container orchestration (EKS/AKS/GKE) * Experience with distributed systems at scale * Familiarity with service meshes and microservices architectures Nice to Have * Experience with chaos engineering tools (Gremlin, Chaos Monkey) * Background in product-facing services with high traffic scale * Understand how to use incident management platforms. This includes using tools like PagerDuty for alerts. It also includes working with DataDog for monitoring. For California, Colorado, Connecticut, Nevada, New York, Rhode Island, and Washington-based candidates: the annual base pay range for this role is listed below. Within this range, individual pay is determined by several factors, including location, job-related skills, work experience, and relevant education and/or training. This role may also be eligible for discretionary bonuses, equity, and/or commissions, as well as benefits. ## Description We're looking for a Senior Software Engineer that will report to the Development Manager / R&D Head. In this role, you will part of the R&D Team that works on mission-critical applications. Your Mission Engage and partner with various Engineering, Operations, and Product teams to design, deliver, and maintain a highly available and performant application platform. * Build and implement application observability and platform monitoring tools to continuously improve the customer experience * Eliminate toil by automating processes, tuning alerts, and improving code where it is most needed * Frequently evaluate new ideas and trends to identify potentially useful tools and techniques * Collaborate with different functional groups to identify gaps, prioritize, and resolve issues * Defining, implementing, and maintaining SLIs and SLOs aligned with customer experience. * Design and instrument SLIs such as latency, error rates, and availability across critical services * Manage and enforce error budgets to balance system reliability with product feature velocity. * Improving alert quality by reducing noise and focusing on actionable, high-signal alerts * Embed with product teams to review architectures and catch reliability risks early * Share your knowledge and experience with the Engineering organization * Share your findings with technical leadership and senior management * Build scripts in python/bash/java or ruby for operational automation and incident response, When you join OneTrust you are stepping onto a launching pad - the countdown has begun. The destination? A career without boundaries working alongside a diverse and inclusive crew who is passionate about doing meaningful work. As a pioneer, your voice and expertise will help chart the direction of an entirely new category. Our commitment to putting people first starts with you. Your growth is part of the mission. Our goal is to give you the power to embark on the next phase of your uniquely, unique career. ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [What Developers Get Wrong About Application Quality](https://www.wearedevelopers.com/videos/233-what-developers-get-wrong-about-application-quality) - [Coffee with Developers: David Heinemeier Hansson](https://www.wearedevelopers.com/videos/875-coffee-with-developers-david-heinemeier-hansson) - [Agentic employees in world's most downloaded FinTech app](https://www.wearedevelopers.com/videos/100123-agentic-employees-in-world-s-most-downloaded-fintech-app) - [NoSQL Data Modeling for Front-end Developers](https://www.wearedevelopers.com/videos/297-nosql-data-modeling-for-front-end-developers) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How Much Does a Software Engineer Make? Realistic Software Engineering Salaries](https://www.wearedevelopers.com/magazine/425-how-much-does-a-software-engineer-make-realistic-software-engineering-salaries) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london) - [Highest Paying Tech Companies in Europe](https://www.wearedevelopers.com/magazine/162-highest-paying-tech-companies-in-europe) - [Are Software Engineer Wages Being Pushed Down? A Report on Tech Salaries](https://www.wearedevelopers.com/magazine/417-are-software-engineer-wages-being-pushed-down-a-report-on-tech-salaries) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk)