DevOps Engineer III
ZoomInfo Technologies LLC
Waltham, MA, United States
about 1 month ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
4 years minimum
Compensation
$113,400.0 - $178,200.0
Working hours
Regular working hours
Job source
Tech stack
Artificial Intelligence
Airflow
Automation of Tests
Big Data
BigQuery
Cloud Computing
Continuous Integration
Data as a Services
DevOps
Data Flow Control
Github
Python (Programming Language)
+25 more
Octopus Deploy
Queueing Systems
RabbitMQ
Reliability Engineering
Prometheus
Workflow Management Systems
Datadog
Data Logging
Scripting
GitHub Copilot
Istio
Grafana
Apache Spark
Event Driven Architecture
Containerization
Git Flow
Kubernetes
Infrastructure Automation Frameworks
Apache Flink
AWS Glue
Performance Monitor
Apache Kafka
Cloudwatch
Terraform
Jenkins
Job description
- Contribute to the design, provisioning, and management of data services on Kubernetes-based platforms (Amazon EKS & Google Kubernetes Engine).
- Implement infrastructure as code with Terraform, ensuring security, scalability, and cost awareness.
- Develop and maintain CI/CD pipelines (Jenkins, Argo CD, GitHub Actions) to enable automated testing and deployments.
- Deploy and support cloud-native data services such as Amazon Kinesis, AWS Glue, Google Pub/Sub, Dataflow, and BigQuery.
- Leverage AI-powered tooling (e.g., GitHub Copilot, generative-AI chat/ops assistants, and AIOps platforms) to accelerate script generation, configuration validation, and incident troubleshooting.
- Create automation scripts and internal tooling in Python to streamline DevOps workflows.
- Assist in establishing monitoring, logging, and alerting using Prometheus, Grafana, CloudWatch, or Datadog; incorporate AI-driven anomaly detection where applicable.
- Participate in on-call rotations, incident triage, and post-incident reviews; apply SRE best practice.
- Collaborate with engineers and software developers to ensure infrastructure aligns with application requirements and company standards.
- Document infrastructure, runbooks, and lessons learned to promote knowledge sharing across teams.
Requirements
- 4-6 years in a DevOps, Site Reliability Engineering, or Cloud Infrastructure role.
- Production experience with AWS and/or GCP data services (e.g., Kinesis, Pub/Sub, Dataflow, BigQuery).
- Hands-on experience managing containerized workloads on Kubernetes (EKS, GKE, or self-managed clusters).
- Solid understanding of Terraform (or similar IaC tools) and Git-based workflows.
- Working knowledge of CI/CD platforms such as Jenkins, Argo CD, and/or GitHub Actions.
- Proficiency with Python or another scripting language for automation.
- Familiarity with observability stacks (CloudWatch, Datadog, etc.).
- Fundamental grasp of SRE principles-service reliability, incident response, and performance monitoring.
- Effective communication skills and a collaborative mindset., * Demonstrated experience using AI-powered copilots, chat assistants, or AIOps platforms to accelerate infrastructure work or incident resolution.
- Experience with workflow orchestration tools (Apache Airflow, Cloud Composer).
- Exposure to big-data frameworks (Spark, Flink) or modern data-lake architectures.
- Knowledge of cost-optimization techniques for cloud resources.
- Familiarity with event-driven architectures and message queues (Kafka, RabbitMQ).
- Understanding of GitOps workflows and service mesh technologies such as Istio.
Benefits & conditions
3.43.4 out of 5 stars Waltham, MA $113,400 - $178,200 a year
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on www.indeed.comGood distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
EM
Eli McGarvie
over 3 years ago
EM
Eli McGarvie
Highest Paying Tech Companies for Developers
over 3 years ago
CH
Chris Heilmann
Dev Digest 121 - AI goes offline
about 2 years ago
CH
Chris Heilmann
Dev Digest 120 - Apple and peers
about 2 years ago
CS
Christina Schaireiter
Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence
2 months ago
CH
Chris Heilmann
Dev Digest 137 - AI'm not sure about this
almost 2 years ago