> Markdown version of [/jobs/ext/2980001-senior-software-engineer-infrastructure](https://www.wearedevelopers.com/jobs/ext/2980001-senior-software-engineer-infrastructure). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Software Engineer, Infrastructure - **Company:** Rad Ai - **Location:** San Francisco, CA, United States (Remote available) - **Experience:** Expert - **Salary:** $160,000.0 - $215,000.0 - **Contract:** Permanent contract - **Skills:** Kubernetes Security, Artificial Intelligence, Amazon Web Services, Amazon Elastic Compute Cloud, Bash Shell, Command-Line Interface, Cloud Computing, Data Stores, Software Debugging, Linux, Apache Hadoop, Python (Programming Language), Linux System Administration, Machine Learning, Networking Basics, Cloudera, Virtual Machines, Data Logging, Pulumi, Grafana, Apache Spark, Hdinsight, Electronic Medical Records, Kubernetes, Infrastructure Automation Frameworks, Health Level Seven International, Terraform, Serverless Computing - **Published:** September 18, 2026 - **Apply:** https://arc.dev/remote-jobs/j/redirect/pl4sjwv0tz ## About the Role * A data-informed approach and a track record of effective problem solving. * Bring 6+ years of hands-on infrastructure / platform development experience (or equivalent practical experience) in modern, cloud-native environments, with a track record of owning critical systems in production. * Extensive experience operating Kubernetes in production, including familiarity with cluster lifecycle management, scaling, and container security. * Proficiency in one or more languages-Python (preferred) and/or Bash-and familiarity with Infrastructure as Code (IaC) tools such as CDK, Terraform or Pulumi. * Strong familiarity with AWS and/or GCP services. * Networking fundamentals and comfort working in command-line Linux environments. * Clear, respectful communication and collaboration skills, including organizing work, delegating when appropriate, and giving/receiving feedback. * Experience designing complex systems and mentoring others in technical design. * Ability to scope project-level work, execute independently, and bring projects to completion while collaborating with teammates. Nice to Haves: * Proficiency in deep Linux troubleshooting, including debugging kernel and driver issues. * Experience in regulated environments (e.g., HIPAA) or at early-stage startups. * Background in healthcare, security, or machine learning. * Familiarity with HL7 or radiology workflows. * Experience with OpenTelemetry or similar tracing services. * Familiarity with Grafana or similar logging services. * Experience with Spark (e.g., EMR, Dataproc, HDInsight) and Hadoop-related technologies. Join our world-class team as we build and deploy AI solutions that empower physicians and transform patient care-making a meaningful impact on millions of lives. Driven by our mission, we prioritize transparency, inclusion, and close collaboration, bringing together exceptional people to revolutionize healthcare. If you're passionate about driving innovation and delivering impactful healthcare solutions, we'd love to hear from you! ## Description The Platform Engineering organization at Rad AI builds the foundations that power all of our products-Reporting, Impressions, and Continuity-and enables product teams to ship reliably, safely, and at scale. Within Platform, the Infrastructure team owns our core cloud infrastructure, platforms, and reliability practices. We're hiring a Senior Software Engineer to help us design and operate robust, scalable systems. In this role, you'll contribute to infrastructure architecture, reliability practices, and thoughtful improvements to our workflows. If you're passionate about building resilient platforms and enjoy collaborating across functions, we'd love to hear from you., * Architect and evolve our cloud infrastructure (primarily on AWS) across container orchestration (Kubernetes, Elastic Container Service), serverless (e.g., Lambda), virtual machines (e.g., EC2), and data stores to support current and future products. * Collaborate with engineering leadership, machine learning, data science, and product partners to help shape our platform vision. * Develop and maintain tooling that improves engineering productivity and developer experience. * Promote sustainable incident response and lead blameless post-incident reviews. * Manage network and systems monitoring, design alert strategies, and participate in an equitable on-call rotation. ## Related Videos - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [Why segmenting your infrastructure into tiers makes your infrastructure design better](https://www.wearedevelopers.com/videos/1960-why-segmenting-your-infrastructure-into-tiers-makes-your-infrastructure-design-better) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [All your telemetry data from any source in one place](https://www.wearedevelopers.com/videos/57-all-your-telemetry-data-from-any-source-in-one-place) - [Unleashing Potential Across Teams: The Power of Infrastructure as Code](https://www.wearedevelopers.com/videos/930-unleashing-potential-across-teams-the-power-of-infrastructure-as-code) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)