AppOpps Engineer

Mphasis
Berkeley Heights, United States
about 1 month ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Working hours
Regular working hours
Job source

Tech stack

Agile Methodology Artificial Intelligence Systems Engineering Computer Programming Continuous Integration Linux DevOps Github Python (Programming Language) Octopus Deploy Reliability Engineering Prometheus
+13 more
Data Logging Google Cloud Cloud Monitoring Grafana Containerization Gitlab-ci Kubernetes Low Latency Deployment Automation Terraform Devsecops Jenkins Artifactory

Job description

  • Google Cloud Platform Infrastructure Management: Design, deploy, and maintain robust infrastructure components, including VPCs, Compute Engine, GKE (Kubernetes), and storage solutions.
  • Automation & IaC: Utilize Terraform or Deployment Manager to manage cloud resources and build CI/CD pipelines to automate deployments. Minimizing manual, repetitive tasks by developing automation scripts and custom tools to streamline deployments and operations.
  • Observability & Incident Management: Develop monitoring, alerting, and logging systems (e.g., Cloud Monitoring, Prometheus, Grafana). Act as primary on-call to troubleshoot production incidents.
  • Incident Management: Serving as a first respon der for system outages and conducting deep-dive root cause analysis (post-mortems) to prevent recurrence
  • CI/CD Pipeline Management: Designing and supporting automated deployment pipelines using Jenkins, ArgoCD, Artifactory, DevSecOps, GitLab CI, or GitHub Actions
  • Reliability Engineering: Define and maintain Service Level Indicators (SLIs) and Service Level Objectives (SLOs) - Latency, Traffic, Errors, and Saturation
  • Functional Domain: Define the pipeline with good understanding of Credit Card processing, Payment transaction
  • Payload: Define automated pipeline applicable to credit card processing for Day-0, Daily basis, and set frequency. Analyse, Triage, Fix and reload payload.
  • Optimization & Security: Proactively optimize infrastructure for cost, performance, and security compliance.
  • Site Reliability Engineer, Google Cloud Engine AI SRE at Google: Focus specifically on AI workload health, and GCE visibility

Requirements

We are seeking a AppOpps Engineer Resource having 8+ years of professional experience ensuring the reliability, scalability, and performance of Google Cloud-based services through automation, monitoring, and proactive engineering. Key responsibilities include managing infrastructure as code (Terraform), optimizing GKE/Kubernetes, incident response, and implementing SLIs/SLOs to minimize manual toil.

This role requires close collaboration with cross functional teams, adherence to DevOps and Agile practices, and ownership of service quality and delivery., * Experience: 8+ years in SRE, DevOps, or systems engineering, specifically with Google Cloud Platform.

  • Technical Skills: Deep knowledge of Linux, Kubernetes (GKE), networking (VPCs, CDNs), and containerization.
  • Programming: Proficiency in scripting/programming languages like Python, Go, or Shell.
  • Methodologies: Strong understanding of GitOps, CI/CD pipelines, and SRE principles (error budgets, toil reduction)
  • Strong troubleshooting skills across the full stack (network, OS, application).
  • Ability to balance system stability with the need for rapid deployment.
  • Observability Tools: Experience implementing monitoring and logging stacks like Prometheus, Grafana, or the Google Cloud Operations Suite
  • Excellent collaboration skills to work with development teams for service ownership

Soft Skills

  • Strong problem-solving and analytical skills
  • Clear communication with technical and non technical stakeholders
  • Ownership mindset and production grade engineering discipline
  • Ability to work independently and within cross functional teams

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

3:32 min

Shifting to a DevOps career from non-technical backgrounds

Megha Kadur · LIVE

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

Videos

See all

Related articles

See all