SRE Engineer

Apptad Inc.
Frisco, TX, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
7 years minimum
Compensation
$107,000.0 - $216,000.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence Amazon Web Services Microsoft Azure Cloud Computing Continuous Integration Cursor (Graphical User Interface Elements) DevOps Python (Programming Language) Performance Tuning RabbitMQ Ansible Azure Machine Learning
+18 more
Shell Script SQL Databases Data Logging Google Cloud System Availability Snowflake Prompt Engineering Reliability of Systems Gitlab Kubernetes Apache Kafka Api Design Terraform Splunk Appdynamics Docker Jenkins Databricks

Job description

Job Description: Note: Fidelity will not provide immigration sponsorship for this position. The Role: As a Principal Engineer on the Enterprise AI/ML Platform team, you will …

  • 1 day ago +

Requirements

  • 5 7+ years of experience as a Senior Systems Reliability Engineer (SRE), DevOps Engineer, or Cloud Infrastructure Engineer in production environments.
  • Strong expertise in CI/CD (GitLab, Azure DevOps, Jenkins), Azure/AWS/GCP, Docker, Kubernetes, Terraform, Helm, Ansible, Python/Shell scripting, SQL, Snowflake, and Databricks.
  • Hands-on experience with monitoring, logging, incident management, root cause analysis (RCA), performance tuning, automation, and infrastructure reliability, along with Kafka/RabbitMQ and Splunk/AppDynamics.
  • Experience leveraging AI-assisted SRE tools such as Claude, Cursor, Prompt Engineering, RAG, and AI Agents to automate incident response, operational workflows, and platform reliability.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:50 min

Navigating specialized roles and toolsets across engineering teams

Nele Uhlemann · WWC 2023

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

Videos

See all

Related articles

See all