Engineer, Site Reliability

T-Mobile Us, Inc.
Frisco, TX, United States
5 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours

Tech stack

Cloud Computing Continuous Integration HP Systems Insight Manager Performance Tuning Reliability Engineering Software Deployment Software Engineering Scripting Cloud Platform System Reliability of Systems Information Technology

Job description

Experteer Overview In this role you will bolster the reliability and resilience of digital infrastructure by automating tasks, monitoring health, and steering incident responses. You’ll work with cross-functional teams to drive uptime and reduce manual interventions, shaping efficient deployment and operational workflows. The position centers on automation, cloud-native practices, and proactive incident management to sustain high service quality. This is an opportunity to influence platform robustness at scale and contribute to a culture of reliability. Compensation / Benefits * Automate processes to improve system reliability and cut manual tasks * Monitor systems proactively to prevent incidents and ensure continuity * Streamline software development and deployment processes for operational efficiency * Develop scripts and tools to reduce routine workloads * Manage incident response for rapid recovery and minimal disruption * Adopt new technologies to sustain system robustness and performance * Participate in additional duties/projects as assigned by management Tasks * Bachelor in Computer Science (or Engineering) with 3 years of related experience; advanced degree with 1 year related experience; or equivalent combination * Experience in CI/CD pipelines for software deployment (preferred) * Experience with cloud-native platforms and solutions (preferred) * Mentoring or guiding teams in reliability engineering practices (preferred) * Knowledge areas: Application Monitoring, Automation, CI/CD, Capacity Planning, Cloud Computing, Incident Management, Performance Tuning, Scripting, System Reliability Key requirements * medical, dental and vision insurance * 401(k) * paid time off and holidays * employee stock grants and ESPP * tuition assistance * mobile service and home internet discounts

Requirements

performance * Participate in additional duties/projects as assigned by management Tasks * Bachelor in Computer Science (or Engineering) with 3 years of related experience; advanced degree with 1 year related experience; or equivalent combination * Experience in CI/CD pipelines for software deployment (preferred) * Experience with cloud-native platforms and solutions (preferred) * Mentoring or guiding teams in reliability engineering practices (preferred) * Knowledge areas: Application Monitoring, Automation, CI/CD, Capacity Planning, Cloud Computing, Incident Management, Performance Tuning, Scripting, System Reliability Key requirements * medical, dental and vision insurance * 401(k) * paid time off and holidays * employee stock grants and ESPP * tuition assistance * mobile service and home internet discounts

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on us.experteer.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

7:01 min

Career progression from backend programming to professional mobile engineering

Edoardo Dusi · LIVE

1:04 min

Introduction to Bitcoin script parsing tools

Steve Shadders · LIVE

2:27 min

Introduction to WebAssembly in a cloud computing context

Edo Edo · WWC 2024

8:32 min

Benchmarking GitOps engine constraints for extensive multi-cluster environments

Artem Lajko · Europe 2026 Virtual

2:48 min

Daily responsibilities and alignment practices for technical engineering leadership

Edoardo Dusi · LIVE

1:53 min

Evaluating traditional scripting languages for modern development tasks

Jens Knipper Jens Knipper · Europe 2026 Virtual

Videos

See all

Related articles

See all