Senior DevOps Engineer
Avantos, Inc.
United States
18 days ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Apply on www.indeed.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Working hours
Regular working hours
Job source
Tech stack
Artificial Intelligence
Amazon Web Services
Amazon S3
Backup Devices
Bash Shell
Cloud Engineering
Continuous Delivery
DevOps
Disaster Recovery
Distributed Systems
Github
Monitoring of Systems
+13 more
Python (Programming Language)
PostgreSQL
Performance Tuning
Reliability Engineering
Datadog
Data Logging
Containerization
Kubernetes
AWS Fargate
Cloudwatch
Amazon Simple Queue Service (SQS)
Terraform
Docker
Job description
We’re seeking a Senior DevOps Engineer / Site Reliability Engineer to own and evolve our infrastructure, reliability, and deployment practices. You’ll be responsible for building the foundational platform that enables our engineering teams to ship quickly and reliably while maintaining the security and compliance standards required in financial services., * Design, implement, and maintain our AWS cloud infrastructure using infrastructure-as-code principles with Terraform
- Build and optimize CI/CD pipelines to enable rapid, safe deployments across multiple environments
- Own observability strategy-implement comprehensive monitoring, logging, and alerting systems using Datadog and other tooling
- Architect and manage containerized workloads on ECS Fargate and evaluate migration paths to Kubernetes
- Establish and enforce security best practices, working closely with compliance teams on financial services requirements
- Design and implement disaster recovery, backup, and business continuity strategies
- Optimize system performance, cost efficiency, and resource utilization across AWS services
- Collaborate with engineering teams to improve service reliability, reduce toil, and establish SLOs/SLIs
- Participate in incident response and conduct thorough post-mortems to drive continuous improvement
- Mentor engineers on DevOps practices, cloud architecture patterns, and operational excellence
Requirements
- 8+ years of experience in DevOps, SRE, or infrastructure engineering roles
- Expert-level proficiency with AWS services including ECS Fargate, ALB, Cognito, S3, SQS, and related services
- Deep hands-on experience with Terraform for managing complex, multi-account AWS environments
- Strong scripting and automation skills in Python and/or Bash
- Proven experience designing and implementing CI/CD pipelines (GitHub Actions, ArgoCD, or similar)
- Solid understanding of containerization technologies (Docker) and orchestration platforms (Kubernetes/ECS)
- Experience with observability and monitoring tools (Datadog, CloudWatch, or equivalent)
- Deep knowledge of networking, security, and AWS best practices
- Strong problem-solving abilities and experience troubleshooting complex distributed systems
- Excellent communication skills and ability to work cross-functionally with engineering teams
Bonus
- Have startup or early-stage experience
- Experience with PostgreSQL performance tuning and RDS management
Why Join
- Build the foundation layer of an AI-native product
- High ownership + direct impact, You’re a systems thinker with a passion for reliability and automation. You love tackling complex infrastructure challenges, thrive in fast-paced environments, and are motivated by the responsibility of building secure, scalable platforms from the ground up. You collaborate deeply, care about operational excellence, and are energized by turning ambiguity into robust solutions that help teams move faster and safer.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on www.indeed.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
EM
Eli McGarvie
over 3 years ago
BR
Benjamin Ruschin
Navigating the AI Shift
11 months ago
IK
Igor Khokhriakov
How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again
3 days ago
CH
Chris Heilmann
Dev Digest 120 - Apple and peers
about 2 years ago
CH
Chris Heilmann
Dev Digest 121 - AI goes offline
about 2 years ago
LM
Luis Minvielle
How to Become an AI Engineer
over 2 years ago