DevOps Engineer

Doctronic Inc.
New York, NY, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$180,000.0 - $240,000.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Amazon Web Services Amazon Elastic Compute Cloud Amazon S3 Bash Shell Cloud Computing Cloud Computing Security Cloud Engineering Continuous Integration Linux DevOps Disaster Recovery
+35 more
Distributed Systems Domain Name System (DNS) Github Identity and Access Management Virtual Private Networks (VPN) Python (Programming Language) Network Security Linux System Administration Machine Learning Routing Octopus Deploy Performance Tuning Reliability Engineering Prometheus Azure Machine Learning Security Information and Event Management Software Deployment Software Vulnerability Management Data Logging Load Balancing Autoscaling Large Language Models Grafana Amazon Virtual Private Cloud (VPC) AI Platforms Gitlab-ci Kubernetes Infrastructure Automation Frameworks Deployment Automation Performance Monitor Hardware Infrastructure Cloudwatch Terraform Devsecops Docker

Job description

We’re looking for a DevOps Engineer to own our infrastructure. We’re HIPAA-compliant and SOC 2 Type II certified - you’ll maintain and strengthen that foundation as we scale to serve millions of patients and enterprise partners.

This role is critical to our mission. When healthcare consultations depend on your infrastructure, reliability isn’t just best practice - it’s a sacred responsibility. You’ll combine hands-on technical work with strategic infrastructure leadership, ensuring Doctronic remains the most trusted AI diagnostic platform in healthcare.

What You’ll Build

Cloud Infrastructure

  • Design, deploy, and maintain AWS infrastructure using ECS and core AWS services, including EC2, IAM, VPC, ALB/NLB, CloudWatch, ECR, S3, RDS, Glue.
  • Operate and scale production Kubernetes clusters with Helm-based application deployments.
  • Implement GitOps workflows with Argo CD to enable secure, automated, and auditable releases.

Infrastructure as Code & CI/CD

  • Provision and manage cloud infrastructure using Terraform.
  • Build and optimize CI/CD pipelines with GitHub Actions, GitLab CI, or similar platforms.
  • Automate infrastructure and operational workflows to improve reliability and reduce manual effort.

AI/ML & LLM Infrastructure

  • Deploy, maintain, and optimize infrastructure for AI/ML services and Large Language Models (LLMs).
  • Support scalable inference workloads with a focus on performance, reliability, and cost efficiency.
  • Collaborate with engineering teams to deliver production-ready AI platforms.

Networking & Security

  • Design and manage cloud networking, including VPCs, VPNs, Load Balancers, DNS, routing, and secure connectivity.
  • Integrate SIEM solutions to improve infrastructure visibility and incident response.
  • Implement security best practices, identity management, and least-privilege access across cloud environments.

Observability & Reliability

  • Build monitoring, logging, and alerting solutions using Prometheus, Grafana, CloudWatch.
  • Improve platform reliability through proactive monitoring, incident response, and performance optimization.
  • Build and optimize containerized workloads using Docker.
  • Administer Linux-based production environments.
  • Automate operational tasks and infrastructure management using Bash and Python.

Requirements

  • 5+ years of experience as a DevOps Engineer or Site Reliability Engineer.
  • Strong hands-on experience with AWS, Kubernetes, Terraform, Docker, Helm, and Argo CD.
  • Proven experience designing and operating highly available, cloud-native production infrastructure.
  • Solid understanding of Linux administration, networking, cloud security, and Infrastructure as Code principles.
  • Experience building and maintaining CI/CD pipelines and deployment automation.
  • Familiarity with monitoring, logging, and observability platforms.
  • Experience supporting AI/ML workloads or modern distributed systems is a strong advantage.
  • Strong problem-solving skills with the ability to troubleshoot complex production environments.
  • Comfortable working in cross-functional teams and collaborating closely with software engineers, security teams, and product stakeholders.
  • Passionate about automation, reliability, scalability, and operational excellence.

Nice to Have

  • Experience with GPU infrastructure and AI/ML model deployment.
  • Familiarity with DevSecOps practices, vulnerability management, and compliance frameworks.
  • Experience with multi-cluster Kubernetes environments.
  • Knowledge of performance tuning, autoscaling, and cost optimization in AWS.
  • Experience supporting high-traffic, mission-critical production systems.
  • Understanding of disaster recovery, backup strategies, and business continuity planning.
  • Experience working in fast-paced startup or scale-up environments.

Benefits & conditions

Pulled from the full job description

  • Health insurance
  • Vision insurance
  • Dental insurance, Base Salary: $180K-$240K + Equity
New York City On-site

Join our NYC team and work directly with engineering and product teams to build security into everything we do.

Equity Opportunities

Share in Doctronic’s growth as we transform healthcare with AI.

Comprehensive Health Benefits

We offer comprehensive health, dental, and vision coverage-plus mental health support and flexible time off-because caring for others starts with caring for ourselves.

Building AI That Matters

Join Doctronic and work with cutting-edge AI that’s transforming healthcare and helping people make faster, smarter decisions.

Reports To

Director of Engineering

Compensation Range: $180K - $240K

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:32 min

Shifting to a DevOps career from non-technical backgrounds

Megha Kadur · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · WWC 2023

1:20 min

Identifying multi-disciplinary talent for developer experience engineering roles

Hazal Mestci +1 · Coffee With Developers

2:39 min

Experiencing core Linux capabilities for DevOps administration

Michael Cade · LIVE

Videos

See all

Related articles

See all