> Markdown version of [/jobs/ext/2295576-sr-site-reliability-engineer-robotaxi-service](https://www.wearedevelopers.com/jobs/ext/2295576-sr-site-reliability-engineer-robotaxi-service). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Sr. Site Reliability Engineer, Robotaxi Service - **Company:** Tesla Motors - **Location:** Austin, TX, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Software Debugging, Linux, Distributed Systems, Domain Name System (DNS), Log Analysis, Networking Basics, Octopus Deploy, Reliability Engineering, Ansible, TCP/IP, Virtual Local Area Networks, Load Balancing, Bug Reporting, Information Technology, Deployment Automation, Terraform - **Published:** August 29, 2026 - **Apply:** https://jobs.localjobnetwork.com/apply/add/88178584/1 ## About the Role * 3+ yearsof SRE, infrastructure engineering, or production operations in a distributed systems environment * Solid Linux and networking fundamentals (TCP/IP, DNS, load balancing, VLAN/MTU) * Hands-onKubernetesin production - pod scheduling, ingress controllers, cluster diagnostics * Observability toolingexperience - log aggregation, metrics, or APM. You build alerts that fire when they should * Proficiency inPython or Gofor scripting, automation, and tooling * Comfort withAI-assisted toolsfor debugging, log analysis, and documentation * Cellular, Starlink, or Wi-Fibehavior at scale in mobile or field-deployed environments * Infrastructure-as-code and GitOps tooling (Ansible, Terraform, ArgoCD, Helm) * Background inIoT, connected vehicle, or fleet management- telemetry pipelines, OTA, vehicle-to-cloud ## Description This role sits within IT Infrastructure, embedded in the Robotaxi operational ecosystem across the US. What You'll Do * Build relationships across Autopilot, fleet networking, platform engineering, and service engineering - becoming the person who understands how the pieces fit and who owns what * Identify where the critical path to reliability and scale is blocked - observability gaps, fragile integrations, processes that break under load - and drive the fix. Translate operational symptoms from the Robotaxi Ops team into well-scoped engineering problems and bring the right teams to a shared solution * Take on the hardest cross-system issues - the ones that don't fit neatly into one team's scope - working through them directly with Autopilot, fleet networking, platform, and infrastructure teams * Use AI-assisted tools for log analysis, anomaly detection, and root cause correlation to reduce time-to-resolution on complex, multi-system issues * Escalate well-scoped problems with clear reproduction steps, supporting data, and impact assessment * Build monitoring, alerting, and dashboards for Robotaxi-critical services - fleet health, tele-operation endpoints, service routing pipelines - integrated with on-call systems * Write automation that reduces toil: fleet health checks, service routing validation, and operational state monitoring * Write runbooks that give Ops and on-call engineers clear, actionable steps when things go wrong * Define priority levels for Robotaxi - full fleet connectivity, tele-operations, data uploads and OTAs. Build the severity classification framework before the incidents arrive * Work closely with Tesla's Incident Management team as the Robotaxi subject matter expert - ensuring alerts route correctly, on-call coverage is scoped appropriately, and postmortem findings feed back into engineering ## Related Videos - [An Applied Introduction to eBPF with Go](https://www.wearedevelopers.com/videos/1075-an-applied-introduction-to-ebpf-with-go) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Dev & Test in the Cloud? Deploy your cloud environments with Ansible & Terraform](https://www.wearedevelopers.com/videos/1607-dev-test-in-the-cloud-deploy-your-cloud-environments-with-ansible-terraform) - [Remote Driving on Plant Grounds with State-of-the-Art Cloud Technologies](https://www.wearedevelopers.com/videos/251-remote-driving-on-plant-grounds-with-state-of-the-art-cloud-technologies) - [Turning Container security up to 11 with Capabilities](https://www.wearedevelopers.com/videos/718-turning-container-security-up-to-11-with-capabilities) - [The best of two worlds - Bringing enterprise-grade Linux to the vehicle](https://www.wearedevelopers.com/videos/67-the-best-of-two-worlds-bringing-enterprise-grade-linux-to-the-vehicle) ## Related Articles - [Dev Digest 137 - AI'm not sure about this](https://www.wearedevelopers.com/magazine/485-dev-digest-137-ai-m-not-sure-about-this) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [The Best X (Twitter) Accounts for Developers](https://www.wearedevelopers.com/magazine/294-the-best-x-twitter-accounts-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs)