> Markdown version of [/jobs/ext/3457965-technology-consultant-site-reliability-engineer-sre](https://www.wearedevelopers.com/jobs/ext/3457965-technology-consultant-site-reliability-engineer-sre). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Technology Consultant - Site Reliability Engineer (SRE) - **Company:** NTT DATA, Inc. - **Location:** Atlanta, GA, United States - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Amazon Web Services, Application Performance Management, Microsoft Azure, Unix, Cloud Computing, Linux, DevOps, Disaster Recovery, Fault Tolerance, Github, Monitoring of Systems, IBM WebSphere MQ, Enterprise Messaging Systems, Performance Tuning, Reliability Engineering, Ansible, Prometheus, Shell Script, Software Deployment, Datadog, Diagnostic Tools, Google Cloud, Enterprise Software Applications, Grafana, Spring-boot, Containerization, Gitlab-ci, Kubernetes, Infrastructure Automation Frameworks, Apache Kafka, Restful APIs, Terraform, Splunk, Dynatrace, Docker, Jenkins, Microservices - **Published:** September 1, 2026 - **Apply:** https://dejobs.org/x/x/85169CE8F4C84135A68D5B129CA94491/job/ ## About the Role * 6+ years of experience in Site Reliability Engineering, DevOps, or Production Engineering/Support. * 4+ years of hands-on experience with Kubernetes, Docker, and containerized application environments. * 4+ years of experience with Java, Spring Boot, Microservices, and REST APIs. * 3+ years of experience with observability and monitoring tools such as Splunk, Dynatrace, Prometheus, Grafana, Datadog, or ELK. Nice to Have * Strong understanding of SLI, SLO, SLA, Error Budgeting, and SRE principles. * Experience with Kubernetes deployment and troubleshooting tools such as Helm. * Experience with AWS, Azure, or Google Cloud Platform. * Knowledge of Linux/Unix and Shell scripting. * Experience with Kafka, IBM MQ, or other messaging technologies. * Knowledge of Terraform, Ansible, or other Infrastructure as Code tools. * Experience with Jenkins, GitLab CI, GitHub Actions, or Azure DevOps. * Experience implementing distributed tracing and application performance monitoring. * Knowledge of incident management and ITIL processes. * Experience supporting high-volume, highly available, distributed enterprise applications. * Strong analytical, troubleshooting, communication, and problem-solving skills. ## Description NTT DATA's Client is currently seeking an experienced Technology Consultant - Site Reliability Engineer (SRE) with strong hands-on expertise in Kubernetes, Observability, Java, and production reliability. The ideal candidate will have experience supporting highly available and distributed enterprise applications, troubleshooting complex production issues, and driving automation and reliability improvements. The role requires close collaboration with application engineering, DevOps, cloud, infrastructure, and support teams to improve application availability, scalability, performance, and operational efficiency. Day to Day Job Duties * Manage and support business-critical applications running on Kubernetes and containerized platforms. * Monitor application and platform health and proactively identify reliability, availability, and performance issues. * Troubleshoot Kubernetes deployments, pods, services, networking, configurations, and application issues. * Implement and enhance observability solutions covering metrics, logs, traces, dashboards, and alerting. * Support and troubleshoot Java/Spring Boot and Microservices-based applications. * Perform root cause analysis (RCA) for critical production incidents and implement permanent corrective actions. * Define and monitor SLIs, SLOs, SLAs, Error Budgets, and other reliability metrics. * Automate repetitive operational activities and identify opportunities to reduce operational TOIL. * Participate in incident, problem, change, and production release management activities. * Collaborate with engineering teams to improve application resilience, performance, scalability, and fault tolerance. * Support CI/CD pipelines and improve application deployment and release processes. * Participate in capacity planning, performance tuning, disaster recovery, and production readiness reviews. * Develop and maintain operational runbooks, troubleshooting procedures, and technical documentation. ## Related Videos - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [SRE Methods In an Agency Environment](https://www.wearedevelopers.com/videos/348-sre-methods-in-an-agency-environment) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [WeAreDevelopers LIVE - Node and Package Security](https://www.wearedevelopers.com/videos/2138-wearedevelopers-live-node-and-package-security) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Résumé-Driven Development: How IT trends affect the job market for software developers](https://www.wearedevelopers.com/magazine/59-resume-driven-development-how-it-trends-affect-the-job-market-for-software-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Find a Developer Job: 12 Best Job Sites For Developers](https://www.wearedevelopers.com/magazine/165-find-a-developer-job-12-best-job-sites-for-developers)